From the Prompt Engineer 48 video. Both models get the same input and typed questions; only the model differs. Laya needs no key.
Catalog, page 2
115 builds · page 2 of 3
One typed question per frame; the other paddle is three lines of arithmetic as a control. Physics in WebAssembly, decisions in native Laya over POST /decide.
The game enumerates placements with their holes, heights and clears; Laya picks one from text. The teacher is a board-evaluation formula, not Jev. Walkthrough video in Japanese.
Deterministic arcade physics, a browser flight recorder and a supervised Laya fine-tune that landed 30 of 30 held-out flights on Apple MLX. 58% less peak memory with 6-bit weights.
GitHubTools & apps
Lunar Laya: Atari-style lander with a CUDA-trained Laya checkpoint that lands 30/30 on MLX
- Held-out landings
- 30 / 30
- Peak memory
- -58% with 6-bit
Native Windows demo with live probabilities, executed moves, inference timing and planner-intervention counts. A deterministic search proposes routes, Laya classifies them.
Unmodified Laya weights, real NES emulation via Stable Retro, text observations from RAM, model-driven jumping. 27.5 s of game time at 16.79 ms median inference.
GitHubTools & apps
Laya clears Super Mario Bros. 3 World 1-1 on a DGX Spark: 207 decisions, frames verified
- Decisions
- 207
- Game time
- 27.50 s
Abhimanyu walks the Mahabharata ring formation to the centre. Each step goes to local Laya as typed questions; unsure or unplayable answers are marked red and corrected.
GitHubTools & apps
Chakravyuha: a browser ring-maze where every move is a Laya decision
Two routers behind one endpoint, measured on 180 labelled requests. Laya answers small/medium/powerful routing in one forward pass, against a GPT-5 nano router.
GitHubLLM routing & guardrails
laya_router: a model router with Laya matches GPT-5 nano's accuracy, 35x faster, free
- vs GPT-5 nano
- same accuracy, 35x faster, as reported
- Requests
- 180
Measures the decision head's attention drop-off on long inputs across 400 documents, then chunks, screens and aggregates, or locates the one passage that matters.
One simulated Swedish villa, four controllers on the same minutes. Laya with a safety clip cut the bill from 985 to 530 SEK at p95 179 ms; a plain tariff EMS reached 385 SEK.
GitHubTools & apps
A public test of Laya on a home energy system: fast enough, but a tariff EMS still wins
- Bill
- 530 SEK vs 985 naive
- Tariff EMS
- 385 SEK
Collects vacancies, filters them against a candidate profile, ranks with the multilingual Laya checkpoint and prints IT and warehouse tables plus a Markdown report.
GitHubDocuments & data
Job search in Denmark: aggregate jobnet.dk and jobindex.dk, rank matches with Laya
Self-hosted console that runs VirusTotal lookups through System One models and evaluates them against each other on a labelled dataset. Prompts and criteria are versioned JSON.
GitHubSecurity & fraud
IOCArena: classify IPs, domains and hashes from VirusTotal data with Jev, Von or Laya
Replayable grid sim with moving workers and forklifts. Laya answers action, collision_risk, path_blocked and needs_operator in one call; collision prevention stays deterministic.
GitHubModeration & safety
Laya Warehouse Safety: a robot picks advance, shift or wait above a deterministic shield
Scores post style locally through a Laya-MLX helper as you scroll, using low-information and formulaic-rhetoric criteria and a conservative consensus score. Abstains when unsure.
GitHubModeration & safety
Slop Finder: a Chromium extension that flags AI-slop style on X, LinkedIn and Reddit
Paste a nervous message and watch tone (choice), formality 1–5 (score) and fight risk (noul) update as you type, 40–110 ms per call. Misses subtle passive-aggression.
GitHubModeration & safety
Before You Send: live tone, formality and fight-risk readings as you type
Companion service that reads Lidarr's manual-import queue, scores every candidate release and track in one pass, and picks the right Single/EP/Album with calibrated probabilities.
GitHubDocuments & data
lidarr-decision-import: resolve Lidarr's stuck imports with Laya instead of an LLM
Describe a past conversation in plain language and Chat Seek finds it across local agent histories, reranks with the receptron Laya runtime, and offers a Resume action.
Allen Porter's native Assist conversation agent scores intents and entities in one Laya pass, runs device commands locally and escalates open-ended queries to a fallback LLM.
GitHubTools & apps
Laya for Home Assistant: a fully local Assist conversation agent
Swift package that reads a form in a running Mac app through the Accessibility API, asks an on-device model what belongs in each field, and types it in. ~1 ms per decision.
GitHubTools & apps
FluidUse: local computer use on Apple Silicon with Laya and CUA-S1-FORMS
Local decider for the browser-agent loop: hand it a numbered table of page controls and it picks the operation and element. Playwright/CDP driver, speaks /v1/systemone.
GitHubLLM routing & guardrails
laya-browser-agent: browser decisions on Laya, drop-in for jev-ultrafast tooling
Openlayer's library packs every eval for a trace into one decision request: tool choice, grounding, scope, injection, PHI. Runs on Jev, or on Kev or Laya locally on a Mac.
GitHubLLM routing & guardrails
jevals: agent evals and guardrails as decisions, runnable on every trace
The decision model never writes code; each step it picks the next tool, scores progress, estimates risk and says if the goal is reached. Any OpenAI-compatible LLM executes.
GitHubLLM routing & guardrails
jeffrey: a coding-agent CLI where Jev or Laya decides and your LLM executes
TypeScript library where one set of question definitions serves a zero-dependency Promise client and an Effect service. Never invents a probability it was not given.
Choice, Score and Noul over any System One backend with a confidence policy: if a backend errors or is unsure, the next is tried. Ships a TypeSafe-compatible server.
GitHubLLM routing & guardrails
pydecide: one Python client and fallback chain for Jev, OpenRouter, Laya, cross-encoders
Hermes plugin that classifies each task against a ~300-skill index with Laya, injects the top skill's SKILL.md as per-turn context, and fails open. Cache-safe by design.
GitHubLLM routing & guardrails
Hermes skill_router: pre-route each task to the right skill with a local Laya model
Cross-platform Agent Skill plus MCP tool that gives coding agents a local Laya decision capability. Verified on Apple Silicon with MLX and PyTorch backends.
SkillLLM routing & guardrails
Laya Router Skill: local decision routing for Codex, Claude Code, OpenCode and Pi
Will Sargent's stdio MCP server on fastmcp. Classify text, score against a rubric or answer yes/no from Claude Code or any MCP client without sending input off the machine.
GitHubTools & apps
laya-mcp: a local MCP server exposing laya-mlx to any MCP client
Homebrew menu bar server that turns n8n Text Classifier requests into Laya typed questions in ~40 ms. Weights ship inside the app; no network, unloads when idle.
GitHubSupport & triage
Laya Serve for macOS: a menu bar app with an OpenAI-compatible endpoint for n8n
- Memory loaded
- ~1.3 GB
Dockerized FastAPI service bundling all three checkpoints behind a language router, with timing-safe API keys, OpenAPI docs, triage and moderation presets and bulk inference.
Typed-decision server on NVIDIA GPUs or Apple Silicon that speaks the Jev wire format, with coding-agent integrations. Built after a cat woke the author at five on a Sunday.
One import, no server. Dependency-free Rust compiled to Wasm with WebGPU kernels; pulls a pinned int8 pack from Hugging Face and keeps it in the browser. Live Tetris included.
Browser runtime for Laya decisions on ONNX Runtime Web. WebGPU with SIMD Wasm fallback, Web Worker friendly, CPU embedding slicing to get past storage-buffer limits.
Live at laya-web.pages.dev. 524 MB of quantized English weights cached in the tab, onnxruntime-web over WebAssembly, and a parity page checked against the PyTorch reference.
Backend-independent Unity client for local Laya servers and TypeSafe-compatible endpoints. Game state in, Choice/Score/Noul out; transport-first, no embedded weights.
GitHubTools & apps
Laya Unity: a Unity-native SDK for System One decisions in games
SwiftPM package ported from laya-coreml. Point it at the aac6fef Core ML bundles (general or ANE) and call predict with choice, score or noul; no Python at runtime.
GitHubTools & apps
LayaKit: a Swift package that runs one Laya decision per call on Core ML
Native Swift library on Apple MLX. Call prepare once to download ~804 MiB of weights, then predict with choice questions straight from your app.
Runs the typed-decisions checkpoint on Metal Performance Shaders: ~32 ms median and ~2.1 GiB RAM on M5 Pro, plus a slower ~0.74 GiB low-memory mode. Demo and benchmark included.
GitHubTools & apps
Laya MPS: typed decisions on a Mac GPU via PyTorch Metal, ~32 ms median
- RAM
- ~2.1 GiB (or ~0.74 GiB mode)
Convert → quantize → deploy path that turns Laya (and Kev, NanoJev, PlayJev) into an INT8 ONNX service on the official wire protocol. Runtime is onnxruntime, tokenizers, numpy.
GitHubTools & apps
EdgeJev: offline Laya on a 4-core CPU, 15.6 ms per question with ONNX INT8
- Jev hosted API
- 314 ms, as reported
One process serves the model API and an admin UI; agents call decide over MCP, HTTP or CLI and take top.id as the action. Windows service installer included.
GitHubTools & apps
laya-go: Go server plus agent CLI and MCP for Laya decisions
Elixir library that downloads the official checkpoint on first load and answers typed choice and noul questions; pick EMLX on Apple Silicon or EXLA for CPU/NVIDIA.
GitHubTools & apps