Google announced Gemini 4 Argon on Tuesday, its newest frontier model built for long-horizon professional work in software engineering, enterprise knowledge tasks, and cybersecurity defense. Writing on the official Google blog, Koray Kavukcuoglu, SVP of Google DeepMind and Chief AI Architect, said the model is rolling out initially to a set of trusted cyber defenders through Google's Fairwind Program, with broader access to follow once guardrails are strengthened.
The announcement lands during one of the busiest stretches of the year for frontier AI releases. For continuous coverage of model launches, policy fights, and funding rounds, AI Buzz Wire tracks the developments that matter as they happen.
A phased rollout shaped by caution
Unlike a typical major model launch, most users cannot try Argon today. Google said it is engaged in the U.S. government's voluntary process for pre-release model access and will gradually expand availability as feedback from early testers shapes its guardrails. Reuters reported the flagship arrives after months of delays, and VentureBeat noted the limited-release strategy still allows Google to claim a retaken benchmark lead over OpenAI and Anthropic. Bloomberg reported measurable skepticism about the new model among Google's own employees ahead of launch.
When wider access comes, Google said it will start with paid API customers and Google AI Ultra subscribers.
What Google says Argon can do
The headline capability change is endurance. Argon's output token limit has been expanded to an industry-leading 1 million tokens, up from the previous 64K. Google argues that giving a model the headroom to reason across hundreds of thousands of tokens in a single trajectory lets it solve hard problems in one pass rather than fragmenting work across sessions.
Google published benchmark results to back the frontier claim:
- DeepSWE v1.1: 77.9% — a new state of the art on the real-world software engineering benchmark
- Vals Index: No. 1 — leading performance across finance, coding, legal, and tax work weighted by contribution to U.S. GDP
- AutomationBench: 51.3% — first place on Zapier's end-to-end business execution benchmark
- LVBench: 91.7% — state of the art in long-video understanding
- CWE-bench v1: 68% — tied for first on security vulnerability remediation
Inside Google's own workflows
Argon is already powering internal work at Google, and the company offered unusually concrete examples. On quantum computing, the model optimized spacetime resources for bottleneck subroutines, beating a published baseline by 40% in minutes. Analyzing fleet-wide profiling telemetry, teams of Argon agents identified and applied memory optimizations across Google's data centers that freed over 300 TiB once rolled out, with an estimated 500 TiB to 1 PiB in total savings.
The model is also driving large-scale code migrations from C/C++ to Rust, scaling from tens of thousands of lines in core libraries like re2 and libgav1 up to more than 800,000 lines in the Fuchsia Zircon kernel. For libgav1, Google's open-source video decoder, Argon agents replaced 32,000 lines of SIMD code with compiler-friendly safe Rust, producing a memory-safe decoder that runs 2.7 times faster than the previous Rust port with identical output.
Cybersecurity defense comes first
Argon was explicitly trained for defensive cyber work: autonomously finding, validating, and patching critical software vulnerabilities. For trusted defenders and Google's internal teams, the company will release the model without cyber guardrails so defenders get full capability. On CWE-bench v1, which measures vulnerability remediation, Argon ties for first place at 68%, building on the frontier performance of 3.8 Flash Cyber, and Google says Argon outperformed that model in attack-surface discovery and proof-of-concept generation on Wiz's internal black-box penetration testing benchmark.
Security firm Wiz is already using Argon through its Scan for Good initiative, which protects critical public infrastructure for free. In an early demonstration, Google says the model uncovered a critical vulnerability exposing sensitive personal information across healthcare software used by hospitals worldwide — a risk previous frontier models had missed.
Four layers of frontier safeguards
وسیع دستیابی سے پہلے، گوگل چار شعبوں میں حفاظتی اقدامات کو مضبوط کر رہا ہے۔ غلط استعمال کے خلاف، آرگن کو دوہری استعمال کی جائز تحقیق کو محفوظ رکھتے ہوئے سائبر اور CBRN حملے کی درخواستوں کو مسترد کرنے کے لیے ڈیزائن کیا گیا ہے، نئی تکنیکوں کے ساتھ جو غلط استعمال کی نشاندہی کرنے کے لیے ماڈل کی اندرونی سرگرمیوں کی نگرانی کرتی ہیں۔ فوری انجیکشن کے خلاف، مخالفانہ تربیت نے آرگن گوگل کا اب تک کا سب سے زیادہ لچکدار ماڈل بنا دیا ہے، جو گرے سوان کے بالواسطہ پرامپٹ انجیکشن کے بینچ مارک پر آگے ہے۔
غلط ترتیب کے لیے، گوگل ایسے تخفیفات کو تعینات کر رہا ہے جو آرگن کی سوچ اور عمل کے سلسلہ کی نگرانی کرتے ہیں، ضرورت پڑنے پر عمل درآمد کو روکتے ہیں، انتباہات کے ساتھ ایک وقف شدہ واقعہ رسپانس ٹیم کو بھیجے جاتے ہیں۔ کمپنی نے اپنی سینڈ باکسڈ ٹریننگ اور ایویلیویشن کے ماحول کو بھی سخت کیا، انہیں زیادہ خطرے سے چلنے سے پہلے سیل کر دیا۔ خاص طور پر، Google نے باقی صنعت پر زور دیا کہ وہ استدلال کی شفافیت کو برقرار رکھیں تاکہ ماڈل کے خیالات غلط ترتیب کی تشخیص کے لیے کارآمد رہیں۔
قیمت اور دستیابی
Argon $2 فی ملین ان پٹ ٹوکنز اور $10 فی ملین آؤٹ پٹ ٹوکن کی تعارفی قیمت پر، کیشڈ ان پٹ ٹوکنز کے ساتھ ان پٹ قیمت کی 95% چھوٹ پر لانچ کرے گا۔ تعارفی مدت کے بعد، قیمتوں کا تعین $4 فی ملین ان پٹ ٹوکن اور $20 فی ملین آؤٹ پٹ ٹوکنز تک بڑھ جاتا ہے - حریف سرحدی پیشکشوں کے خلاف جارحانہ انداز میں Argon کی پوزیشننگ۔
اسٹیجڈ ڈیبیو فرنٹیئر ریلیز کے لیے ایک نئی حقیقت کی عکاسی کرتا ہے: صلاحیت اب ٹیبل اسٹیک ہے، اور فرق کرنے والا کنٹرول کا مظاہرہ کر رہا ہے۔ گوگل شرط لگا رہا ہے کہ سائبر ڈیفنڈرز، ہسپتال کے سافٹ ویئر، اور زنگ کی منتقلی صحیح نمائش ہیں - اور یہ کہ ایک سست عوامی رول آؤٹ کی قیمت کسی اور حفاظتی واقعے سے کم ہے۔ ڈویلپرز اور انٹرپرائزز کو ادا شدہ API اور AI الٹرا ایکسیس ونڈو کو دیکھنا چاہیے، جسے گوگل کا کہنا ہے کہ جیسے ہی گارڈریل ٹیسٹنگ کی اجازت دی جائے گی کھل جائے گی۔
---
AI سے آگے رہیںتازہ ترین AI خبریں، تجزیہ اور کامیابیاں حاصل کریں — سب ایک جگہ پر۔
مزید AI خبریں پڑھیں →