Target: Provider-neutral model profilesBaseline: Raw, prefix, and fixed-ratio baselines
v0.6 Adaptive Context Planner
Frozen multilingual context planning with explicit quality and safety checks
The offline suite compares raw context, naive prefix truncation, fixed-ratio compression, and adaptive planning. It uses task-grounded string constraints rather than an LLM judge.
Empirical benchmark matrix
Limitations and non-recommended workloads
- The corpus is compact and synthetic; it is not a universal task-quality claim.
- Latency is environment-specific and the checked-in run is not a throughput guarantee.
- The live Sarvam harness requires explicit opt-in and credentials and is not part of the offline result.