February 2026
Model
NGen 4 Lite
A fast, efficient NGen 4 reasoning model for real-world product interactions.

Overview
NGen 4 Lite is a fast and efficient reasoning model optimized for real-world interactions. It gives teams a lightweight NGen 4 option with reliable response quality and practical production latency.
Context & Specs
Model Priority
Lightweight reasoning
Mode
Reasoning-capable Lite tier
Best For
High-volume assistants, internal tools, quick reasoning tasks, and cost-aware production systems.
System Card
NGen 4 system card linked above.
Key Capabilities
- Fast and reliable reasoning for daily AI workflows.
- Efficient profile for product teams scaling NGen 4 usage.
- Strong instruction following and multilingual coverage for the Lite tier.
- Benchmark coverage across reasoning, long context, coding, agent, and multilingual tasks.
Selected release
NGen 4 Lite
1 internal variant
Reasoning
Medium
Speed
High
Input price
Rs 24 / 1M input tokens
Output price
Rs 40 / 1M output tokens
Input
Text, image, audio, video
Output
Text, image, audio, video
Overview
Model capabilities, context, and selection details.
NGen 4 Lite has a 256K context window with a High speed tier. Reasoning tier: Medium. Input support: text, image, audio, video. Output support: text, image, audio, video.
- Context window: 256K (262,144 tokens).
- Speed tier: High.
- Reasoning tier: Medium.
- Input support: text, image, audio, video. Output support: text, image, audio, video.
Context window
256K (262,144 tokens)
Max output
64K (65,536 tokens)
Knowledge cutoff
Jan 2025
Rate limit
1,200 RPM
Pricing
Pricing is based on the selected internal release.
Input
Rs 24 / 1M input tokens
Cached input
Rs 6 / 1M cached input tokens
Output
Rs 40 / 1M output tokens
Model Evaluations
NGen 4 Lite Evaluations
ShadCN benchmark charts for NGen 4 Lite using the benchmark sheet you provided, with Qwen 3.5 4B and 9B removed.
NGen 4 Lite Core Reasoning & Knowledge
Advanced knowledge, academic reasoning, and evaluation performance across core reasoning benchmarks.
NGen 4 Lite Instruction Following
Instruction adherence and challenge-following capability across structured user intent tasks.
NGen 4 Lite Long Context
Retention and retrieval performance over longer context windows.
NGen 4 Lite Reasoning & Coding
Contest math and code generation benchmarks for the Lite tier.
NGen 4 Lite General Agent
Tool use, planning, and general agent benchmark performance.
NGen 4 Lite Multilingual
Cross-lingual reasoning, translation, and multilingual benchmark coverage.

Build with TNSA models
Explore model families, snapshots, and developer tools for production AI systems.