Rabbit Industries public reference • updated 2026-09-30

Agastya Competition Reference

A detailed specification and benchmark reference for Agastya AGI v01 alongside selected frontier systems. The position column is Rabbit Industries' internal research/reference ordering; it is not a third-party global leaderboard. Agastya's external same-suite performance rank remains pending frozen evaluation and independent replay.

Rank-format correction: every model now uses the same #1–#7 reference-position format. Agastya is displayed as #7, with a separate PUBLIC CANDIDATE / external rank pending status.
Reference positionModel / providerAccess / architectureTotal paramsActive paramsContextTraining tokensPublic benchmark evidenceQualification statusPrimary source
#1GPT-6 AstraOpenAI • Sep 3, 2026REFERENCE LEADERAccessClosed frontierReasoning • coding • computer use • researchTotalUndisclosedActiveUndisclosedInput1,050,000128K max outputPretrainingUndisclosedARC-AGI-3 99.9%GPQA Diamond 96.0%Terminal-Bench 4.0 57.9%DeepSWE v1.1 74.1%HLE w/tools 57.2%PUBLIC FRONTIERSame-suite Agastya parity not yet tested.OpenAI official
#2Claude Opus 5.5Anthropic • Sep 22, 2026FRONTIER AGENTICAccessClosed frontierAgentic coding • knowledge work • long-running tasksTotalUndisclosedActiveUndisclosedStandard paid/API200K+Provider/platform limits can varyPretrainingUndisclosedTerminal-Bench 4.0 66.4%FrontierCode 1.1 Main 54.4%HLE w/tools 67.7%GDPval-AA v2.1 1846 EloPUBLIC FRONTIERSame-suite Agastya parity not yet tested.Anthropic official
#3Gemini 3.8 FlashGoogle • Sep 2026 stableFRONTIER MULTIMODALAccessClosed frontierText • image • video • audio • PDFTotalUndisclosedActiveUndisclosedInput1,048,57665,536 max outputPretrainingUndisclosedGPQA Diamond 95.3%DeepSWE v1.1 73.8%Terminal-Bench 4.0 19.1%PUBLIC FRONTIERGoogle official
#4Qwen3.7-MaxQwen / Alibaba • May 20, 2026FRONTIER AGENTAccessProprietaryLong-horizon agent foundationTotalUndisclosedActiveUndisclosedInput1,000,000PretrainingUndisclosedGPQA Diamond 92.4%HLE 41.4%HLE w/tools 53.5%Terminal-Bench 2.0 69.7%TB 2.0 is not TB 4.0.PUBLIC FRONTIERQwen official
#5DeepSeek-V4.1-FlashDeepSeek • Sep 10, 2026FRONTIER AVAILABLEArchitecture552B MoECausal Encoder–Decoder • native multimodalTotal552BActive8B input / 16B outputContext1,000,000PretrainingUndisclosedGPQA Diamond 90.9%HLE 36.8%HLE w/tools 63.9%Terminal-Bench 4.0 31.2%DeepSWE v1.1 74.2%PUBLIC FRONTIERDeepSeek official
#6Llama 4 MaverickMeta • Apr 5, 2025OPEN-WEIGHT REFERENCEArchitectureMoE multimodal128 routed experts + shared expertTotal400BActive17BContext1M classScout is Meta's 10M-context variantFamily mixture>30TLlama 4 family mixtureOfficial Meta release publishes broad multimodal, reasoning and coding comparisons.Not placed into the 2026 same-suite matrix where methods differ.OPEN-WEIGHTMeta official
#7Agastya AGI v01Rabbit Industries Pvt Ltd • Bramha Medha Native INDIAPUBLIC CANDIDATEArchitectureDense nativeSovereign random-init Native Brain lineageTotal1.265BExact: 1,264,715,776Active1.265BExact: 1,264,715,776 • 100% activeContext4,096Canonical context tokensPretraining tokens1.000BExact snapshot: 1,000,255,070Comparable public scores: PENDINGBM-1B acceptance: pendingIndependent replay: pendingFrozen same-suite frontier results: not yet publishedPUBLIC CANDIDATEReference position #7 • external/global performance rank not established.Rabbit Industries official profilePublic specifications + qualification statusNative Brain Evidence R1

Same-benchmark evidence matrix

Benchmark versions are kept explicit. A dash means this page is not asserting a directly verified value from the cited official source. Agastya stays pending rather than receiving an estimated score.

BenchmarkGPT-6 AstraClaude Opus 5.5Gemini 3.8 FlashQwen3.7-MaxDeepSeek V4.1 FlashLlama 4 MaverickAgastya AGI v01
GPQA Diamond96.0%—95.3%92.4%90.9%—PENDING SAME-SUITE RUN
Terminal-Bench 4.057.9%66.4%19.1%— (69.7% is TB 2.0)31.2%—PENDING SAME-SUITE RUN
DeepSWE v1.174.1%—73.8%—74.2%—PENDING SAME-SUITE RUN
HLE w/tools57.2%67.7%—53.5%63.9%—PENDING SAME-SUITE RUN

Agastya disclosure — aligned with the same comparison format

Reference position#7

Same # format as the other entries; separate external-rank status avoids ambiguity.

Parameters1.265B total / active

Exact: 1,264,715,776 total and active parameters. Dense model, 100% active.

Canonical context4,096 tokens

Current canonical context; experimental long-context work is tracked separately.

Pretraining tokens1.000B tokens

Exact current evidence snapshot: 1,000,255,070 tokens.

Native lineageRandom initialization

Canonical lineage does not use third-party model weights as initialization.

Benchmark stateQualification in progress

No same-suite frontier score is published until frozen evaluation and independent replay complete.

External rankNot established

External leaderboard acceptance is a future evidence gate.

Company designationThe First World Native Environment

Rabbit Industries designation; independent global-first verification remains pending.

Methodology: Reference positions are Rabbit Industries' research/reference ordering, not a universal intelligence score. Vendor results can differ by harness, effort level, tool access, safeguards and benchmark version. A direct Agastya performance rank requires identical frozen suites and independent replay.