Tagged articles

model pricing

16 articles · Page 1 of 1
Top Architect
Top Architect
Sep 5, 2026 · Artificial Intelligence

Google's Triple Gemini Launch: 3.6 Flash, Flash-Lite, Cyber & Gemini 4 Pre-training Begins

Google DeepMind releases three specialized Gemini models — 3.6 Flash for token-efficient reasoning, 3.5 Flash-Lite for high-speed low-cost volume tasks, and 3.5 Flash Cyber for vulnerability detection — while confirming aggressive pre-training for Gemini 4, signaling a push to make production AI agents faster, cheaper, and more capable.

AI agentsGeminiGoogle DeepMind
0 likes · 7 min read
Google's Triple Gemini Launch: 3.6 Flash, Flash-Lite, Cyber & Gemini 4 Pre-training Begins
Top Architecture Tech Stack
Top Architecture Tech Stack
Aug 22, 2026 · Artificial Intelligence

GPT‑5.6 Sol price cut cuts model spend by 20% – developers need to recalc costs

With the GPT‑5.6 Sol API and token pricing reduced by over 20% for the next three months, teams must reassess unit‑task costs, adopt multi‑layer optimization—request tiering, context management, agent round‑control, and caching—to decide when the flagship model is truly cost‑effective.

AI agentsContext ManagementGPT-5.6
0 likes · 10 min read
GPT‑5.6 Sol price cut cuts model spend by 20% – developers need to recalc costs
AI Code to Success
AI Code to Success
Aug 14, 2026 · Artificial Intelligence

DeepSeek’s Double Launch: V4‑Pro Model Goes Live and Harness Open‑Source, Advancing Agents

On August 13, DeepSeek simultaneously released the flagship V4‑Pro model and open‑sourced its Harness runtime, illustrating the “Agent = Model + Harness” paradigm; the article details the model’s pricing, performance, new features, the plugin‑centric design of Harness, usage steps, community ecosystem, and broader AI‑agent implications.

AI AgentDeepSeekHarness
0 likes · 11 min read
DeepSeek’s Double Launch: V4‑Pro Model Goes Live and Harness Open‑Source, Advancing Agents
Design Hub
Design Hub
Aug 14, 2026 · Artificial Intelligence

Four AI Releases in One Day: What’s Shaping the Emerging AI Delivery Stack?

On a single day, Google, DeepSeek, and MiniMax unveiled Gemini 3.7 Flash, V4‑Pro, the Harness runtime, and Music 3, each illustrating how AI is shifting from headline‑grabbing benchmarks toward cost‑effective agents, plug‑in runtimes, and controllable content generation for real‑world workflows.

AI agentsDeepSeek HarnessDeepSeek V4 Pro
0 likes · 13 min read
Four AI Releases in One Day: What’s Shaping the Emerging AI Delivery Stack?
Coder Life Journal
Coder Life Journal
Jul 10, 2026 · Artificial Intelligence

OpenAI Unveils GPT-5.6 with Sol, Terra, Luna Tiers and Codex Merged into ChatGPT

On July 9, OpenAI released GPT‑5.6, introducing three tiered models—Sol, Terra, and Luna—with distinct pricing and performance, merging the standalone Codex app into ChatGPT, adding an ultra‑mode for multi‑agent collaboration, enhanced design capabilities, and showing benchmark gains over competitors.

AI benchmarksChatGPTCodex
0 likes · 4 min read
OpenAI Unveils GPT-5.6 with Sol, Terra, Luna Tiers and Codex Merged into ChatGPT
AI Insight Log
AI Insight Log
Jul 8, 2026 · Artificial Intelligence

Cursor and SpaceXAI Unveil Grok 4.5: Musk Says It Matches Opus 4.7 Performance

Cursor teams up with SpaceXAI to launch Grok 4.5, a model trained on real development data and engineered tasks, whose benchmark scores approach GPT‑5.5 and Opus 4.8, while Musk claims its speed rivals Opus 4.7, and the model is now priced for subscription use across desktop, web, iOS, CLI and SDK.

AI benchmarksCursorGrok 4.5
0 likes · 6 min read
Cursor and SpaceXAI Unveil Grok 4.5: Musk Says It Matches Opus 4.7 Performance
DataFunTalk
DataFunTalk
Jun 27, 2026 · Artificial Intelligence

OpenAI Unveils GPT‑5.6 ‘Solar System’ Models: Sol, Terra, Luna Outperform Mythos

OpenAI released GPT‑5.6 with three tiered models—Sol, Terra and Luna—named after celestial bodies, offering lower pricing, record‑breaking benchmark scores in programming, security, biology and health, new max and ultra inference modes, limited partner access, and a deployment plan on Cerebras that could make it the fastest flagship LLM.

AI benchmarksGPT-5.6Large Language Model
0 likes · 8 min read
OpenAI Unveils GPT‑5.6 ‘Solar System’ Models: Sol, Terra, Luna Outperform Mythos
AI Insight Log
AI Insight Log
Jun 27, 2026 · Artificial Intelligence

GPT-5.6 Crushes Claude Fable 5 in TerminalBench – What This Means for AI

OpenAI's GPT-5.6 tops the TerminalBench 2.1 leaderboard, introduces a three‑tier model line (Sol, Terra, Luna) with aggressive pricing, limits access despite strong security‑focused benchmarks, and signals a shift toward tiered, infrastructure‑style releases for high‑capability AI models.

AI model benchmarkingAI safetyClaude Fable 5
0 likes · 8 min read
GPT-5.6 Crushes Claude Fable 5 in TerminalBench – What This Means for AI
AI Engineering
AI Engineering
May 28, 2026 · Artificial Intelligence

Anthropic Unveils Claude Opus 4.8: Same Price, Agent Power Beats GPT‑5.5

Anthropic released Claude Opus 4.8 with unchanged pricing, new inference‑strength controls, Dynamic Workflows for massive tasks, a fast mode 2.5× quicker and three‑times cheaper, and benchmark results showing its agent capabilities surpass GPT‑5.5 while improving honesty and alignment.

AI agentsAnthropicClaude Opus 4.8
0 likes · 12 min read
Anthropic Unveils Claude Opus 4.8: Same Price, Agent Power Beats GPT‑5.5
AI Large Model Application Practice
AI Large Model Application Practice
Apr 24, 2026 · Artificial Intelligence

DeepSeek V4 Preview: Key Technical Highlights, Benchmarks, and Pricing

The DeepSeek‑V4 preview details two model variants—Pro and Flash—with trillion‑scale parameters, outlines benchmark scores that surpass or match leading overseas models across code generation, real‑world fixes, engineering tasks, and world knowledge, and explains core innovations, pricing, API endpoints, and open‑source licensing.

APIDeepSeekLLM
0 likes · 7 min read
DeepSeek V4 Preview: Key Technical Highlights, Benchmarks, and Pricing
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Feb 20, 2026 · Artificial Intelligence

Google Reclaims AI Throne with Gemini 3.1 Pro, Achieving 77.1% ARC‑AGI‑2 Score

Google’s Gemini 3.1 Pro, the latest upgrade to the Gemini 3 series, achieves a verified 77.1 % score on the ARC‑AGI‑2 reasoning benchmark—more than double the performance of Gemini 3 Pro—while leading in GPQA, LiveCodeBench Pro, SWE‑Bench Verified, and MMMLU tests, and is now being rolled out to developers, enterprises and consumers with detailed pricing and integration options.

AI benchmarkingARC-AGI-2Gemini 3.1 Pro
0 likes · 9 min read
Google Reclaims AI Throne with Gemini 3.1 Pro, Achieving 77.1% ARC‑AGI‑2 Score
Wuming AI
Wuming AI
Feb 20, 2026 · Artificial Intelligence

Gemini 3.1 Pro: How Google Boosted Reasoning Scores and What It Means for Developers

Google's Gemini 3.1 Pro preview raises reasoning benchmark scores dramatically, offers new pricing tiers, and is already integrated into Gemini API, CLI, Vertex AI, and consumer apps, while community demos showcase SVG animation, real‑time dashboards, 3D simulations, and heat‑transfer analysis.

AI benchmarksGemini 3.1 ProGoogle AI
0 likes · 5 min read
Gemini 3.1 Pro: How Google Boosted Reasoning Scores and What It Means for Developers
AI Insight Log
AI Insight Log
Dec 11, 2025 · Artificial Intelligence

GPT-5.2 Released: How It Outperforms Claude 4.5 and Gemini 3 Pro

OpenAI’s GPT‑5.2 launch introduces three specialized modes, achieves a record 55.6% score on SWE‑Bench Pro, demonstrates strong front‑end generation, adds a /compact API for long‑context efficiency, offers tiered pricing with cache discounts, and improves safety for younger users.

AI benchmarkingAI safetyFront-end generation
0 likes · 6 min read
GPT-5.2 Released: How It Outperforms Claude 4.5 and Gemini 3 Pro
AI Algorithm Path
AI Algorithm Path
Jun 11, 2025 · Artificial Intelligence

OpenAI's O3‑Pro Model: Deep Reasoning, Pricing, Benchmarks, and Access Guide

OpenAI introduced the O3‑Pro multimodal deep‑reasoning model with an 80% price cut for O3, detailed its training via large‑scale reinforcement learning, compared its capabilities and costs against GPT‑4o, GPT‑4.1 and O3‑Pro, listed its core specs, limitations, access methods, and presented benchmark tests that highlight both strengths and weaknesses.

AIO3-ProOpenAI
0 likes · 10 min read
OpenAI's O3‑Pro Model: Deep Reasoning, Pricing, Benchmarks, and Access Guide