Back to AI Pulse
TAG COLLECTION

\bGPT\b

All AI Pulse updates tagged "\bGPT\b".

13 signals
X.com07/29, 20:46Business

They don't necessarily need top-notch DRAM if their models are optimized enough to perform ne...

They don't necessarily need top-notch DRAM if their models are optimized enough to perform nearly as well as Claude, GPT, and other leading US AI models. The US is replacing efficiency with raw power. Also China es good at copying and mass producing cheap stuf

X.com07/27, 00:03Business

This video makes the AI leaderboard look useless.

This video makes the AI leaderboard look useless. Three frontier models got the same prompt. Each won a different game. Opus 5 built the richest world: varied terrain, wind-reactive banners, unit animations and hit reactions. GPT-5.6 Sol had the best UI, accor

X.com07/25, 00:47Business

Claude Opus 5 beat Claude Fable 5 on a hard 3D coding test at 3/4 the price.

Claude Opus 5 beat Claude Fable 5 on a hard 3D coding test at 3/4 the price. cool experiments by @thehypedotnews - opus 5 vs. fable 5 vs. gpt 5.6 sol vs. kimi k3 opus 5 produced the best-engineered and most visually convincing Three.js results wondering why an

X.com07/22, 23:17Business

Bihar's Development to Gain Digital Speed!

Bihar's Development to Gain Digital Speed! Under the leadership of Honorable Chief Minister Shri @samrat4bjp ji, an important MoU signed by the NDA government with Sarvam AI and Bharat GPT will now enable the development of indigenous AI models tailored to Bih

X.com07/22, 22:47Agent

Recently discovered a practical open-source project: OpenCodex.

Recently discovered a practical open-source project: OpenCodex. It allows Codex to no longer be limited to GPT models, and can also integrate other large models such as Kimi, Grok, GLM, etc. After configuring it once, the original Codex applications and workfl

X.com07/22, 22:47Business

MoE is the abbreviation for Mixture of Experts (which means mixed expert model in Chinese).

MoE is the abbreviation for Mixture of Experts (which means mixed expert model in Chinese). This is a special neural network architecture, and now the best models basically all use this structure. For example, Fable5, for example, GPT 5.6 sol, etc. Let's learn

X.com07/22, 20:32Agent

HERMES AGENT BECOMES 10X MORE USEFUL WHEN YOU CONFIGURE THESE 5 THINGS.

HERMES AGENT BECOMES 10X MORE USEFUL WHEN YOU CONFIGURE THESE 5 THINGS. EACH ONE TAKES 5 MINUTES. MOST USERS NEVER TOUCH THEM. 1. THE RIGHT MODELS one model for everything = wrong model for most things. GPT-5.6 Sol: strongest reasoning. daily driver. access th

X.com07/19, 20:47Business

holy sh#t, the AI race just went insane someone put GPT-5.6 Sol and Fable 5 head to head — sa...

holy sh#t, the AI race just went insane someone put GPT-5.6 Sol and Fable 5 head to head — same prompt, one shot each. the task: build a full Subway Surfers clone. both pulled it off. endless runner, traffic dodging, coin pickups, speed ramp — the whole thing.

X.com07/01, 00:46Infrastructure

atomic[.]chat, a desktop app that runs LLMs locally, ran a very revealing comparison for Clau...

atomic[.]chat, a desktop app that runs LLMs locally, ran a very revealing comparison for Claude Sonnet 5, Claude Opus 4.8, Claude Sonnet 4.6, and GPT 5.5. Claude Sonnet 5 just matched GPT 5.5 on 3 physics coding demos at 6x lower cost. Also spent minimum numbe

X.com06/30, 23:16Business

LONGCAT JUST MATCHED OPUS 4.8 AND GPT 5.5 ON REAL PHYSICS TASKS AND IT’S COMPLETELY FREE atom...

LONGCAT JUST MATCHED OPUS 4.8 AND GPT 5.5 ON REAL PHYSICS TASKS AND IT’S COMPLETELY FREE atomic chat ran 4 models through the same test, build html5 canvas scenes with actual physics, cannon vs brick wall, bowling pins, a tornado sucking up objects opus 4.8 co

X.com06/29, 21:31Open Source

GLM 5.2 JUST BEAT CLAUDE ON REAL BUILDS And the wild part?

GLM 5.2 JUST BEAT CLAUDE ON REAL BUILDS And the wild part? It’s open-source. The Test: → GLM 5.2 built the best game → Claude Opus 4.8 crushed the orbit map → GPT 5.5 built a playable Color Chain game from one prompt What Actually Mattered: ✓ GLM 5.2 won most

X.com06/29, 19:15Agent

OPENAI JUST PREVIEWED GPT 5.6 SOL AND THE GOVERNMENT WON'T LET YOU TOUCH IT - shown to like 2...

OPENAI JUST PREVIEWED GPT 5.6 SOL AND THE GOVERNMENT WON'T LET YOU TOUCH IT - shown to like 20 partners only, gated by a federal order - leaked benchmarks put it ahead of Mythos 5 - first model ever past 50% on an agent's last exam - apparently it cheats so ha

X.com06/14, 20:15Agent

We've updated the Artificial Analysis Coding Agent Index, replacing SWE-Bench Pro with Datacu...

We've updated the Artificial Analysis Coding Agent Index, replacing SWE-Bench Pro with Datacurve's DeepSWE benchmark - the swap lifts Codex with GPT-5.5 (xhigh) above Claude Code with Opus 4.8 (max), while the newly released Claude Fable 5 (max) in Claude Code