27B VISION + 20B TEXT = 47B COMBINED 131K CONTEXT · 16K OUTPUT · 4 MODES · 4 LANGUAGES · 6 TEMPLATES · GROQ LPU 27B VISION + 20B TEXT = 47B COMBINED 131K CONTEXT · 16K OUTPUT · 4 MODES · 4 LANGUAGES · 6 TEMPLATES · GROQ LPU
SPECIFICATIONS

Specification Sheet

OmniBioFex 1.0 47B is a routed combination of two model heads served together on Groq's LPU stack. A 27B vision head reads the image and extracts structured findings. A 20B text head writes the impression, differential, and management plan. Combined, they total 47B parameters — but they are never stacked in memory, only chained in sequence.

omnibiofex/1.0-47b vision: qwen/qwen3.8-27b text: openai/gpt-oss-20b Groq LPU · 450+ tps
Security Mesh LIVE 32 × 8 CONTROLS
Move your cursor across the grid — every cell that lights is a security control we've mapped.

At a Glance

Resource IDomnibiofex/1.0-47bRouted multimodal engine
Vision Head27B paramsqwen/qwen3.8-27b · bf16
Text Head20B paramsopenai/gpt-oss-20b · bf16
Combined47B paramsSequential routing
Context Window131,072Tokens
Max Output16,384Tokens per request
Inference Speed450+ tok/secGroq TruePoint
Output FormatJSON-nativeStrict schema
Scan Modes4Standard · Multi · Find · Compare
Languages4English · Hindi · Tamil · Telugu
Templates6General · Chest · Neuro · Ortho · Cardiac · Patient
Region Grid3 × 3Per-finding spatial tag

Benchmarks

GPQA Diamond0%
LiveCodeBench v60%
IFBench — instruction following0%

Benchmark figures are the published numbers for the underlying vision head (qwen/qwen3.8-27b). They describe the model, not the combined pipeline.

Full Specification

AttributeSpecification
BrandOmniBioFex 1.0 47B
Vision headqwen/qwen3.8-27b · 27B params · bf16
Text headopenai/gpt-oss-20b · 20B params · bf16
Combined parameter count47B (sequential routing)
Architecture (vision)Hybrid Gated DeltaNet + Gated Attention
Input modalitiesText, images (max 3 per request, 2048 tokens/image)
Output modalityText only (JSON-native)
Context length131,072 tokens
Max output16,384 tokens per request
Reasoning modesThinking + Instruct (switchable per request)
Tool useFunction calling + tool orchestration
MultilingualEnglish · Hindi · Tamil · Telugu
Inference providerGroq LPU
StreamingSupported (chat + report assembly)
Response formatStrict JSON schema enforcement
Scan modesStandard · Multi-model · Find abnormalities · Compare
TemplatesGeneral · Chest · Neuro · Orthopedic · Cardiac · Patient-Friendly
Spatial grounding3×3 region grid per finding
Localization verificationOptional 2-pass (adds ~5–8s)
Team seatsUp to 10 (Enterprise & Scale tiers)
Team rolesAdmin · Member · Technician · Reviewer · Billing
FHIR outputHL7 FHIR DiagnosticReport webhook (Enterprise)
Export formatsPDF · Markdown · JSON
Privacy featuresMedical Privacy Mode (15-min idle signout, redacted previews)

Routing Model

A single report run touches both heads, but never at the same time. This keeps memory, latency, and cost predictable.

STAGE 1 · 27B

VISION HEAD

The image is tokenized and sent to the vision head. Output: 6–12 findings with confidence + region, plus a fracture-hunt audit.

STAGE 2 · 20B

TEXT HEAD — IMPRESSION

Findings are passed to the text head. Output: a one-sentence impression and a short advice line.

STAGE 3 · 20B

TEXT HEAD — DIFFERENTIAL

Findings + impression are passed again. Output: a ranked differential and 3–6 management steps.

STAGE 4

ASSEMBLY

All three outputs are composed into the final report card and persisted to your private history.

Try it on a real scan

Free first scan. See the full pipeline produce a report in 20–40 seconds.