Back to AI Pulse
TAG COLLECTION

model

All AI Pulse updates tagged "model".

47 signals
X.com08/09, 00:08Agent

Alibaba's ABSeeker adds step-level credit assignment to long-horizon search agents It backtra...

Alibaba's ABSeeker adds step-level credit assignment to long-horizon search agents It backtracks from the answer to recover clues, then scores each step, rewarding useful actions in failed trajectories and suppressing errors in successful ones. 4B model matche

X.com08/08, 23:32Business

The great silence of no more high energy particles is going to continue.

The great silence of no more high energy particles is going to continue. Colliders are cool but the Standard Model is all there is likely to be, just need to find that right handed neutrino in a big vat.

X.com08/08, 22:47Business

Donald Trump bought a brand new Tesla Model S Plaid in Ultra Red at full price as show of sup...

Donald Trump bought a brand new Tesla Model S Plaid in Ultra Red at full price as show of support in light of all the Tesla hate recently. From not being invited to the EV Summit to having the whole Tesla lineup parked in front of the White House.

X.com08/08, 22:03Business

Everyone's interpreting the increasingly jargon-y ways of frontier models as a a fuckup in th...

Everyone's interpreting the increasingly jargon-y ways of frontier models as a a fuckup in their training. Seems obvious to me that "confused" is just what it feels like to be talking to an intelligence greater than one's own. Many of us are just getting a fee

X.com08/08, 19:01Open Source

I'm releasing Roomform, an open-source alternative to the RoomPlan API.

I'm releasing Roomform, an open-source alternative to the RoomPlan API. It ingests point cloud scans and extracts geometry including walls, doors, windows, objects. No rectangular-room or Manhattan layout assumptions. The model also infers structure behind occ

X.com08/02, 23:46Business

Workbuddy's edge lies in the fact that Tencent's own large model isn't strong.

Workbuddy's edge lies in the fact that Tencent's own large model isn't strong. Tencent isn't pushing this Harness as a large model company, so from day one, it had to be model agnostic. Once you build a model-agnostic Harness, you break free from the infightin

X.com08/02, 23:17Business

This statement is correct.

This statement is correct. Without multimodality, your model can only be one of the optional models and cannot become the main model.

X.com08/02, 22:17Business

ten significant advances in mathematics and theoretical computer science.

ten significant advances in mathematics and theoretical computer science. solved using an internal version of Astra, our next major model, for a total cost of about $2000 at Sol API prices:

X.com08/02, 22:04Business

I love dropping floor plans of houses in the Bay Area and asking LLMs to redesign in a differ...

I love dropping floor plans of houses in the Bay Area and asking LLMs to redesign in a different aesthetic, Kyoto in this case. Models have gotten phenomenal at 3D.

X.com07/27, 00:03Business

If you let a general-purpose large model hand-write Three.js code to build a character 3D mod...

If you let a general-purpose large model hand-write Three.js code to build a character 3D model, it's like using a missile to swat a mosquito! I did exactly that last week. An Odyssey character—tweaked it back and forth for three hours, burned through hundreds

X.com07/25, 00:47Business

Opus 5 crushed Fable 5 at 3D destruction physics for 2x cheaper!

Opus 5 crushed Fable 5 at 3D destruction physics for 2x cheaper! We gave four models the same task: build three self-contained HTML scenes with real physics Prompts: - A tornado that sucks in a whole field - A wrecking ball taking down an apartment block - An

X.com07/21, 22:17Business

World Labs founder @drfeifei is acquiring SceniX, the robotics simulation team built by @Yunz...

World Labs founder @drfeifei is acquiring SceniX, the robotics simulation team built by @YunzhuLiYZ . World models were already about 3D space. This pushes the work closer to robot training, where simulated worlds have to survive contact with real hardware.

X.com07/20, 21:02Business

Patrick Boyle, after 20+ years in quantitative trading: “The job is not to trust your model.

Patrick Boyle, after 20+ years in quantitative trading: “The job is not to trust your model. It is to keep trying to prove it wrong.” In 8 minutes he explains how quants turn market opinions into hypotheses, test them against historical data, calculate costs a

X.com07/19, 20:47Business

A TEAM STOPPED PAYING FOR THEIR BEST MODEL ON EVERY SINGLE REQUEST Most teams either lock int...

A TEAM STOPPED PAYING FOR THEIR BEST MODEL ON EVERY SINGLE REQUEST Most teams either lock into one expensive model for everything or bounce between providers with no real logic behind it. An engineer set up routing instead, sending routine scripts to a cheap m

X.com07/02, 00:31Safety

FABLE 5 IS BACK.

FABLE 5 IS BACK. AND IT'S ABOUT TO DISAPPOINT EVERYONE WHO WAITED. Two weeks of hype. Two weeks of "the best model is coming home." Here's what you're actually getting. A new safety filter tighter than anything they've shipped before. Normal coding and debuggi

X.com07/01, 21:33Business

you think you checked the answer, because you asked it "are you sure?" you didn't the model c...

you think you checked the answer, because you asked it "are you sure?" you didn't the model can't feel the difference between knowing and guessing, so that question just makes it apologize, hand you another version, wrong in a new way you tested nothing you nu

X.com06/25, 21:46Product Update

Customer by customer it is gonna take an entire year to release the model.

Customer by customer it is gonna take an entire year to release the model. The fact that government is micromanaging, the release is insane.

X.com06/25, 19:15Open Source

“The #3 closed-source LLMs is most in trouble.

“The #3 closed-source LLMs is most in trouble. As enterprises adopt model routing, they're increasingly choosing between the top proprietary models and rapidly improving open-source alternatives. That leaves the number three closed-source provider squeezed fro

X.com06/24, 20:30Research

This diagram shows the paper’s idea that intelligence can be modeled as a small meaning-based...

This diagram shows the paper’s idea that intelligence can be modeled as a small meaning-based decision map, where clues like danger, food, urgency, and location push the system toward 1 of 4 actions: hide, escape, stay still, or move slowly. The point is that

X.com06/24, 19:04Infrastructure

Two Transcription Models in Parallel

Two transcription models run in parallel, one excels in quality, the other in silence detection.

X.com06/18, 23:33Agent

Excited to announce the first workshop on Learning from Situated and Embodied Interaction @ #...

Excited to announce the first workshop on Learning from Situated and Embodied Interaction @ #COLM2026! What can interaction with environments, humans, and other agents teach language models that passive text cannot? … arning-situated-interaction.github.io Subm

X.com06/18, 22:18Business

Just to be clear, if you remove Fable which is unavaialble, GLM-5.2 (Max) is the #1 model in ...

Just to be clear, if you remove Fable which is unavaialble, GLM-5.2 (Max) is the #1 model in the world for frontend coding. This is a huge moment. OSS has caught up with proprietary, and China has caught up with the US, in this very important domain.

X.com06/17, 22:30Business

Car owners are showing off Tesla's overseas version of the Grok large language model combined...

Car owners are showing off Tesla's overseas version of the Grok large language model combined with FSD, leaving many domestic car companies speechless! @Tesla_AI @Grok @Tesla

X.com06/17, 20:20Model Release

Introducing GLM-5.2, Our Latest Flagship Model

GLM-5.2 marks a significant leap in long-horizon task capability.

X.com06/17, 19:45Infrastructure

Deep Agents Deep Dive Part 3 | Delegation

A planning tool that helps models organize work for challenging tasks.

X.com06/15, 20:15Research

Creating low-latency text-to-speech models

It's hard to make models with sub-300ms median ttft.

X.com06/13, 21:30Research

Coding is a clear step up from glm-5.1

Coding shows significant improvements in context and memory.

X.com06/12, 20:45Research

BraceSproul and jakebroekhuizen discuss open source models

They share insights on open source models.

X.com06/10, 20:31Product Update

Meet DiffusionGemma!

Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Moving beyond sequential, token-by-token processes to generate entire blocks of text simultaneously. Here’s what’s new with

X.com06/10, 00:15Business

Whoa, I might’ve just experienced the Longest Continuous Tesla Actually Smart Summon ever in ...

Whoa, I might’ve just experienced the Longest Continuous Tesla Actually Smart Summon ever in my Model 3. My car drove itself for 0.4 miles for over 2:40 straight without stopping once. This was epic!

X.com06/09, 19:17Model Release

Highly Anticipated Mythos Model Just Dropped

The highly anticipated Mythos model has been released, priced significantly higher than previous models.

X.com06/07, 23:15Business

Here's a teaser of our Mac-1 model.

Here's a teaser of our Mac-1 model. 6.6B model, runs locally (on any Mac), requires 7GB RAM (12GB ideal), can use 487 MacOS native tools, perform multi-tool chained tasks, reasoning: ON, output: ~65 tok/s. We built a robust application layer around the model t

X.com06/07, 20:30Product Update

Cheapest high-performance long-context model

Nemotron 3 Ultra from NVIDIA is the cheapest high-performance long-context model.

X.com06/07, 02:04Research

Competition math won't be interesting soon

"Pretty soon, competition math, competition coding, is not going to be interesting anymore."

X.com06/05, 17:02Business

Option pricing models to become free by 2026

Option pricing models that cost $50k/year will be free by 2026.

X.com06/04, 17:00Product Update

New course on serving LLMs efficiently

Learn how to serve models to many concurrent users at low latency.

X.com06/04, 12:30Infrastructure

Chinese Student Runs High-Cost Models with One Chip

A Chinese student runs $1,900/month models using just one chip.

X.com06/04, 08:30Research

Discussing Sakana AI Project on TV Tokyo Tonight

Tonight, discussing Sakana AI's 1T parameter model project on TV Tokyo.