Open proof
84%

MedQA on Chaperone-Thinking-LQ-1.0 — within 4 points of GPT-4o, ~20GB, 1.6× faster than the 32B base.

Models

Chaperone-Thinking-LQ-1.0

Open reasoning model: GPTQ + QLoRA on DeepSeek-R1-Distill-Qwen-32B. Medical and scientific corpora. Fully available on Hugging Face.

View results →
Chaperone-Coder-LQ-1.0

Quantized coding assistant for debugging and production snippets. Same deployability story: small enough for a single L40/L40s.

Hugging Face →

Applications

How the language line shows up for teams who never want to touch a checkpoint.

Chatbots

Conversational assistants on your data, privately hosted. Support teams stay on hard tickets; the model handles the rest.

Explore chatbots →
Specificity

Semantic analysis on specialist text — trends, sentiment, and extraction powered by Thinking-LQ, not a generic API.

Explore specificity →

Different corpus, private deploy, or another language task?

Book a technical call