Personal Project · Sep 2026
Laya Decision Studio
Local Multimodal Typed-Decision Workbench
Overview
Built a local FastAPI workbench for typed yes/no, choice and score decisions over text, images and video frames, with an evaluation harness that separates working software from model quality.
Highlights
- 01
Served the Laya decision model and Qwen3-VL-2B vision locally with pinned model revisions, explicit device reporting and bounded image and timestamped-frame inputs, plus labeled evaluation that reports accuracy, Brier score, NLL and ECE.
- 02
Ran an independent 24-request Coding Lab benchmark in which Laya got more individual decisions right than a rules baseline (238/288 vs. 226/288) but fewer complete specifications (2/24 vs. 5/24), and reported that result as measured.
- 03
Connected the model to Codex through an MCP server and CLI with a versioned repair ledger, and kept it out of review filtering after a held-out diagnostic caught only 1 of 12 positive cases.
Stack
- Python
- FastAPI
- PyTorch
- Hugging Face Transformers
- Qwen3-VL-2B
- SmolVLM2
- Laya
- Model Context Protocol (MCP)
- Pytest
- Model Calibration
- CUDA
Working on something similar?
Tell me about the problem and the data behind it. I reply within 24–48 hours.
Discuss a project