‹ Back to Home

OpenAI Debuts GPT-5.3 Codex: The First AI That Helped Build Itself — What It Means for Developers, India, and the Future of Coding

OpenAI has unveiled GPT-5.3 Codex, the first AI model that was instrumental in creating itself. Running 25% faster than its predecessor, scoring record benchmarks, and raising unprecedented cybersecurity concerns — here is everything you need to know about the most capable coding AI ever released.

Keerthika 7 min read 1,521
Follow on Google
Updated 2 weeks ago
AI Tools OpenAI Debuts GPT-5.3 Codex: The First AI That Helped Build Itself — What It Means for Developers, India, and the Future of Coding 7 min left Follow on Google
OpenAI Debuts GPT-5.3 Codex: The First AI That Helped Build Itself — What It Means for Developers, India, and the Future of Coding

TamilTech AI summary

OpenAI launched GPT-5.3 Codex on February 5, 2026, as the first production model that helped build itself—engineers used early versions to debug training runs, manage GPU clusters, and diagnose evaluations, though it remains a human-guided tool rather than fully autonomous self-improvement. It arrived just 20 minutes after Anthropic’s Claude Opus 4.6, runs about 25% faster with fewer tokens than GPT-5.2 Codex, and can work across millions of tokens via compaction on tasks far beyond coding, including apps, PRDs, research, and desktop work. Benchmarks show strong gains—77.3% on Terminal-Bench 2.0, 64.7% on OSWorld, 70.9% on GDPval, and 77.6% on cybersecurity CTFs—while Claude Opus 4.6 still leads deep debugging on SWE-Bench Verified at 80.8%. That cyber “High capability” rating worries experts, so OpenAI is delaying full API rollout, adding identity checks for advanced cyber features, and funding defense research. Indian users can get it on ChatGPT Plus at ₹1,999/month or Pro at ₹19,900/month, alongside a Codex desktop app past 500,000 downloads and the new Frontier agent platform, making it a major speed boost for developers if they weigh the real security risks as the AI race accelerates.

  • What does "self-improving" mean for GPT-5.3 Codex?
  • How much does GPT-5.3 Codex cost in India?
  • Is GPT-5.3 Codex better than Claude Opus 4.6?
  • Why is OpenAI delaying API access for GPT-5.3 Codex?

AI-assisted summary, checked by the TamilTech editorial team.

0:00
0:00
🔒 Listen is for subscribers. Subscribe

On February 5, 2026, OpenAI launched GPT-5.3 Codex — a model that represents a watershed moment in artificial intelligence history. For the first time, an AI model was instrumental in creating itself. Early versions of GPT-5.3 Codex were used by OpenAI engineers to debug training runs, manage GPU cluster deployments, and diagnose complex evaluation results during the model's own development.

This is not science fiction. This is not a theoretical paper. This is a production AI model available today to millions of ChatGPT users worldwide, including in India at ₹1,999/month (Plus) and ₹19,900/month (Pro).

In perhaps the most dramatic launch timing in AI history, GPT-5.3 Codex was announced just 20 minutes after Anthropic dropped its own flagship model, Claude Opus 4.6. The result? The most intense head-to-head AI showdown we've ever witnessed.

What Makes GPT-5.3 Codex Different From Everything Before

GPT-5.3 Codex is not just an incremental upgrade. It fundamentally changes the relationship between AI and its own development process. Here's what's new:

1. Self-Improving: The AI That Built Itself

OpenAI's official blog states: "GPT-5.3-Codex is our first model that was instrumental in creating itself." Specifically, the Codex engineering team used early versions of the model to:

  • Debug its own training code — finding and fixing bugs in the training pipeline
  • Manage GPU cluster deployments — handling infrastructure during traffic surges
  • Diagnose evaluation results — interpreting complex test outcomes

Sam Altman himself confirmed: "It was amazing to watch how much faster we were able to ship 5.3-Codex by using 5.3-Codex, and for sure this is a sign of things to come."

Important nuance: This is NOT fully autonomous recursive self-improvement. The model was used as a tool by human engineers within a tightly controlled lab environment. OpenAI's own internal assessment states GPT-5.3 Codex does not reach "High capability" on AI self-improvement. Think of it as a brilliant intern that accelerated its own team's work — not Skynet building itself.

2. 25% Faster, Fewer Tokens

Compared to GPT-5.2 Codex, the new model runs 25% faster while using fewer tokens. This means lower costs for API users and snappier responses for ChatGPT subscribers. For developers running large codebases through the model, this efficiency gain translates directly to savings.

3. Beyond Just Coding

While called "Codex," this model goes far beyond writing code. OpenAI claims it can do "nearly anything developers and professionals can do on a computer":

  • Build complex web-based applications during multi-day sessions
  • Debug, deploy, and monitor production applications
  • Write product requirement documents (PRDs)
  • Edit copy, conduct user research, create presentations
  • Handle accounting spreadsheets, scheduling, and diagrams

4. Millions of Tokens in a Single Task

GPT-5.3 Codex is natively trained to operate across multiple context windows through a technique called "compaction." This allows it to coherently work over millions of tokens in a single task — processing entire codebases, not just individual files.

Benchmark Performance: The Numbers Tell a Story

Here's how GPT-5.3 Codex performs against its predecessors and competitors:

BenchmarkGPT-5.3 CodexGPT-5.2 CodexClaude Opus 4.6Human
SWE-Bench Pro (real-world coding, 4 languages)56.8%56.4%——
Terminal-Bench 2.0 (CLI/agentic coding)77.3%64.0%65.4%—
OSWorld-Verified (desktop productivity)64.7%38.2%—~72%
GDPval (44 occupations, knowledge work)70.9%———
Cybersecurity CTF77.6%———
SWE-Bench Verified (bug fixing)——80.8%—

Key Takeaways from Benchmarks

  • Terminal-Bench 2.0: GPT-5.3 Codex at 77.3% absolutely crushed the competition — a 13-point leap over its predecessor and 12 points ahead of Claude Opus 4.6
  • OSWorld: Score nearly doubled from 38.2% to 64.7%, approaching human-level performance (72%)
  • SWE-Bench Verified: Claude Opus 4.6 leads at 80.8%, suggesting Anthropic's model may be better at deep debugging tasks
  • GDPval: Scoring 70.9% across 44 different occupations shows this isn't just a coding tool — it's a general knowledge worker

GPT-5.3 Codex vs Claude Opus 4.6: The 20-Minute Showdown

In what tech media is calling "the most dramatic AI launch week ever," OpenAI and Anthropic released their flagship models within 20 minutes of each other on February 5, 2026. Both companies also aired competing Super Bowl advertisements on February 9.

FeatureGPT-5.3 CodexClaude Opus 4.6
ReleaseFeb 5, ~7:00 PMFeb 5, ~6:40 PM
StrengthSpeed, CLI tasks, structured outputDeep reasoning, debugging, long sessions
Terminal-Bench77.3% (winner)65.4%
SWE-Bench Verified—80.8% (winner)
Key InnovationSelf-improving / built itselfAdaptive Thinking (adjusts reasoning depth)
ContextMillions of tokens (compaction)1 million tokens (native)
Unique FeatureCybersecurity vulnerability detectionAgent teams (parallel agents)

Developer testing reveals nuanced differences. In one debugging test, GPT-5.3 Codex "ran more than eight forensic tool calls but missed the actual problem," while Opus 4.6 "read the document structure once and diagnosed the issue." This suggests Codex excels at breadth and speed, while Opus excels at depth and precision.

The Cybersecurity Elephant in the Room

Here's the part that has security experts worried. GPT-5.3 Codex is OpenAI's first model classified as "High capability" for cybersecurity under their Preparedness Framework. Fortune's headline reads: "OpenAI's new model leaps ahead in coding capabilities — but raises unprecedented cybersecurity risks."

Why Is This Concerning?

  • The model was directly trained to identify software vulnerabilities
  • It scored 77.6% on Cybersecurity CTF challenges — professional-level performance
  • OpenAI admits it "could meaningfully enable real-world cyber harm, especially if automated or used at scale"

What OpenAI Is Doing About It

  • Trusted Access for Cyber — first-ever pre-access identity verification for a cyber-capable AI
  • Delayed API access — full API is being rolled out cautiously, not all at once
  • $10 million in API credits for cybersecurity defense research
  • Safety training, automated monitoring, and enforcement pipelines
  • Advanced cyber functions require identity verification

India Pricing and Availability

GPT-5.3 Codex is available to Indian users through ChatGPT's paid plans with local INR pricing — India is one of the few countries outside Europe to get this:

PlanPrice (India)GPT-5.3 Codex Access
ChatGPT Free₹0No
ChatGPT Go₹399/monthNo
ChatGPT Plus₹1,999/monthYes
ChatGPT Pro₹19,900/monthYes (priority)

What This Means for Indian Developers

  • At ₹1,999/month, Indian developers get access to the most capable coding AI ever built — that's less than ₹67/day
  • For India's massive IT services industry (TCS, Infosys, Wipro), this could dramatically accelerate development velocity
  • The cybersecurity capabilities at 77.6% are directly relevant for India's growing cybersecurity sector
  • GDPval covering 44 occupations means impact extends beyond pure coding into business operations
  • The ₹19,900/month Pro tier is steep for individual developers but a no-brainer for companies and startups

The Codex Desktop App: 500,000+ Downloads

Alongside the model, OpenAI's Codex desktop app has crossed 500,000 downloads. The app provides:

  • Direct access to GPT-5.3 Codex through a native interface
  • IDE integration via extensions
  • CLI (command-line interface) for terminal-based workflows
  • Multi-day session support for complex projects

Frontier Agent Management Platform

OpenAI simultaneously launched the Frontier agent management platform:

  • Natural language interface for building AI agents
  • Integration with CRM platforms, data warehouses, and enterprise services
  • User-created "skills" to extend agent functionality
  • Memory systems for agents to improve over time
  • Dashboard monitoring with performance metrics
  • Limited enterprise access initially (Oracle and HP as launch partners)

The Bigger Picture: Where Are We Heading?

Sam Altman stated in October 2025 that OpenAI aimed for "an automated AI research intern by September 2026" and "a true automated AI researcher by March 2028." GPT-5.3 Codex is a clear step in that direction.

Anthropic's CEO Dario Amodei echoed this trajectory: "We essentially have Claude designing the next version of Claude itself... that loop starts to close very fast."

The enterprise AI market is shifting. OpenAI's wallet share among enterprises has declined from 62% in 2024 to a projected 53% in 2026, while Anthropic has grown from 14% to 18% in the same period. The competition is intensifying, and developers are the ultimate beneficiaries.

Should You Use GPT-5.3 Codex? Our Take

Best For:

  • Terminal/CLI-heavy workflows — 77.3% Terminal-Bench score is unmatched
  • Fast iteration on large codebases — 25% speed improvement matters
  • Full-stack development — goes beyond coding into deployment, monitoring, documentation
  • Cybersecurity research — if you're in infosec, the CTF capabilities are extraordinary

Consider Alternatives If:

  • You need deep debugging precision — Claude Opus 4.6 leads in SWE-Bench Verified at 80.8%
  • You need complex reasoning — Opus 4.6's Adaptive Thinking may be more suitable
  • You're budget-constrained — Anthropic and Google offer competitive free tiers

Conclusion

GPT-5.3 Codex represents a milestone in AI development — the first model to meaningfully participate in its own creation. While "self-improving AI" sounds alarming, the reality is more nuanced: it's a powerful tool that accelerated human engineers' work, not an autonomous agent redesigning itself.

For Indian developers, at ₹1,999/month, it offers extraordinary value. The 77.3% Terminal-Bench score, 64.7% OSWorld performance approaching human levels, and the ability to process millions of tokens in a single task make it the most capable coding assistant available today.

But the cybersecurity concerns are real. OpenAI's decision to delay API access and require identity verification for advanced cyber features shows they're taking the risks seriously. The question isn't whether AI models will keep getting more powerful — it's whether our safety frameworks can keep pace.

The AI arms race just entered a new phase. And it launched in 20 minutes.

Get tomorrow’s tech news on WhatsApp

One short update a day, free. Follow the TamilTech channel.

What do you think?

people reacted

Keerthika

TamilTech editorial team · 3,344 articles

Keerthika is an editor at TamilTech, the Tamil and English technology publication founded by Praveen Kumar S. She covers AI, smartphones, gadgets, EVs, startups and cybersecurity i...

More from Keerthika

Ask TamilTech on WhatsApp

Tech doubt? Ask in Tamil or English — our WhatsApp assistant answers from TamilTech articles in seconds.

Related stories

Comments (0)

| Supports **bold**, *italic*, `code`

Be the first to comment!

Next story Claude Opus 5.5 Tested: What's New and How Good Is It, Really?
Tamiltech

Tamiltech

Install app for faster access

Earn XP 🏆
WhatsApp
Notifications