TNSA

November 2025

Chat Model

NGen 3.5 Lite

Efficient conversational intelligence for high-volume product and support workflows.

NGen 3.5 Lite

Overview

NGen 3.5 Lite provides efficient and fast conversational AI capabilities, optimized for applications that need quick response times, low operating cost, and steady everyday reasoning quality.

Context & Specs

Context Length

128,000 tokens

Model Priority

Low-latency chat

Best For

Support flows, product assistants, high-volume messaging, and lightweight content tasks.

Pricing

Rs 0.30 input / Rs 0.45 output per 1K tokens

Key Capabilities

  • Fast and efficient conversational model for repeated production use.
  • Optimized for quick response times across common chat and support workloads.
  • Cost-effective profile for high-volume applications.
  • Balanced performance across knowledge, reasoning, and instruction-following tasks.

Performance Benchmarks

CategoryBenchmarkScoreIndustry Avg.
KnowledgeMMLU-Pro70.481.2
KnowledgeMMLU-Redux83.7-
ReasoningGPQA55.968.4
ReasoningAIME 2565.670.9
TNSA platform

Build with TNSA models

Explore model families, snapshots, and developer tools for production AI systems.