← Dhi Labs B1 · PROMPT2MODEL

Prompt2Model

A compiler for vision models: describe the detector you want in plain English, get a deployed TensorRT engine, or a documented refusal, not a silent degradation.

Research prototype (source not published)

What it does

Prompt2Model takes a plain-English description of a detector you want and runs it through a full pipeline: prompt parsing, dataset loading (classification and COCO-style detection), training, ONNX export, and evaluation.

An opt-in factory-compiler layer adds an LLM planner, a distill-then-quantize-then-accuracy-floor compression gate, pluggable deployment targets, calibration and conformal abstain behavior, and a flywheel hard-case store.

The classification path is described in its own README as fully validated end to end. The detection path is integrated through the project's week-4 milestone but is not yet fully validated at every stage, and the README says so directly.

Evidence

Synthetic

There is no accuracy or latency benchmark table in the README yet. The demonstrated mechanism is the compression refusal gate: it refuses to ship a compressed artifact that falls below max(accuracy_floor_relative * baseline_accuracy, constraints.accuracy_floor) (default relative floor 0.98), and keeps the uncompressed model instead. This is a behavioral guarantee, not a measured accuracy number. The README states the classification path runs end to end on synthetic, generated toy data.

Test count

Measured 2026-09-26 in a fresh venv: 232 passed, 0 errors, 0 failures.

Demo

Limits

Reproducibility

The source is not published. An outsider can try the hosted demo Space and request research access to the code through dhi-tech.com. What cannot be reproduced today: a real-image accuracy benchmark, because none has been published yet, and full validation of the detection path, which the project's own README flags as not yet complete.