Cinematic Industrial Facility
MEND - X Β· MULTI-MODEL INFERENCE MATRIX

Three Intelligence Tiers.
Zero Hallucination.

One generic LLM cannot solve factory downtime. PLCs demand sub-100ms edge speed; complex breakdowns require deep deductive reasoning. MEND-X dynamically routes every query to the exact intelligence tier needed.

< 100ms
Minimum Latency
100%
Manual Grounded
3 Tiers
Dynamic Routing
0.0%
Hallucination Target
TIER 00 Β· RESILIENT MULTI-PROVIDEREDGE DEPLOYABLE
Auto Round-Robin

Dynamic Failover: Ollama Cloud β†’ Local β†’ Groq

Our most resilient inference setup. Dispatches queries to Ollama Cloud first for high throughput. If network drops or quotas exhaust, it gracefully cascades to your local Ollama daemon, followed by Groq as emergency backup.

Underlying LLM EngineOllama Cloud API / Local Ollama / Groq Fallback
Average LatencyAdaptive (~400ms Cloud / ~2s Local)
Context Window128,000 tokens
Inference Throughput300+ tokens/sec (Cloud) / 45 tokens/sec (Local)
Recommended HostCloud GPU + Local Edge Hybrid

Engineered Superpowers

Zero downtime with automatic 3-tier round-robin fallback
Prioritizes fast cloud GPU inference when API key is present
Offline continuity with local Ollama runtime
Strips chain-of-thought tokens cleanly for valid JSON output

Target Real-World Queries

> Complex multi-machine error code diagnosis under fluctuating connectivityAuto Round-Robin Handled
> Heavy maintenance procedure synthesis with zero downtime requirementAuto Round-Robin Handled
> Offline-first industrial troubleshooting on factory floor gatewaysAuto Round-Robin Handled
> Automated sensor telemetry and fault code correlationAuto Round-Robin Handled
LIVE INFERENCE ROUTER

How MEND-X Decides the Tier

Click a real maintenance scenario below to see the heuristic complexity analyzer evaluate the query and dynamically activate the optimal model.

ROUTER_TRACE // CLASSIFIER_V2
REALTIME ANALYSIS
INPUT_PROMPT: "PowerFlex 755 Fault 8: Step-by-step deceleration profile tuning and motor test"
TARGET_MACHINE: Allen-Bradley PowerFlex 755
COMPLEXITY_METRIC: 0.58 / 1.00(Procedural Repair)
DECISION_RATIONALE: Multi-step mechanical maintenance procedure requiring sequential action items, tool specs, and parameter verification.
ROUTED_LLM: Groq LPU / openai/gpt-oss-20bLATENCY: 1.0s – 1.8s
TECHNICAL SPECIFICATIONS

Side-by-Side Comparison

Metric / CapabilityNord (Tier 01)Forge (Tier 02)Apex (Tier 03)
Primary ObjectiveSub-100ms Error Code TriageMulti-Step Repair SequencesRoot Cause & Safety Critical
Base LLM Enginegroq/compound-mini (Groq LPU)openai/gpt-oss-20b (Groq LPU)openai/gpt-oss-120b (Groq LPU)
Response Latency< 100ms1.0s – 1.8s2.0s – 3.8s
Context Window8,192 tokens128,000 tokens128,000 tokens
Edge / Offline CapableYes (Local IPC / ONNX)Yes (Plant Server)Air-Gapped Private VPC
Hallucination MitigationStrict RAG MaskingPage Citation GroundingRefusal Circuit + Thresholds
Trigger Thresholdcomplexity < 0.350.35 ≀ complexity < 0.70complexity β‰₯ 0.70

Experience the Inference Routing in Real-Time

Test how MEND-X queries live OEM manuals and streams citation-verified repair protocols to line operators in under 8 seconds.

MEND-XMEND-X v1.2.1PROD
DIMENSITY LABS [VH26-37]β€’VCET NATIONAL HACKATHON 2026β€’From Failure to Function
Β© 2026 MEND-X. All rights reserved.