Back to AI Pulse
TAG COLLECTION

Infrastructure

All AI Pulse updates tagged "Infrastructure".

45 signals
X.com08/01, 23:46Infrastructure

Managing a crypto portfolio shouldn't be complicated.

Managing a crypto portfolio shouldn't be complicated. That's why I'm checking out @alloxdotai AlloX uses AI to individual tokens, users can choose a automical A smart way to stay diversified and invest with data-driven insights. #Allox #Crypto #aigirlep1 #Web3

X.com07/31, 21:33Infrastructure

Disagree, its still AI slop.

Disagree, its still AI slop. Completely derivative, but visually interesting at a first glance. Spending $100s in tokens to copy Minecraft or No Man's Sky or random FPS is mind blowing technically, but its not imaginative. Real creativity from AI is coming but

X.com07/29, 22:32Infrastructure

The strongest monetary anchor is whatever the global economy must use to function at the time.

The strongest monetary anchor is whatever the global economy must use to function at the time. Now that is a full stack economy doing things at the necessary cost of production not gold, oil, AI input/output tokens.

X.com07/29, 00:02Infrastructure

jackie szymanski @SzymanskiJackie · 3h inc.com Move Over, Unlimited PTO: The New Must-Have AI...

jackie szymanski @SzymanskiJackie · 3h inc.com Move Over, Unlimited PTO: The New Must-Have AI Perk Taking Over Silicon Valley Nvidia CEO Jensen Huang proposes offering engineers token budgets for AI use. 3

X.com07/27, 22:46Infrastructure

The stocks are down but our Tesla future in space has never been more likely nor more imminen...

The stocks are down but our Tesla future in space has never been more likely nor more imminent than it is today. Tesla ability to immediately create and power the new MegaPod nodes for a grid improving, distributed AI compute cluster for inference questions on

X.com07/27, 20:16Infrastructure

THE AI RACE HAS A BOTTLENECK It’s not GPUs.

THE AI RACE HAS A BOTTLENECK It’s not GPUs. It’s POWER. $KEEL is positioning itself around the infrastructure layer of AI with data center projects and a massive power pipeline. The crowd is watching the chip makers… Smart money watches what the chips NEED.

X.com07/27, 00:47Infrastructure

PEOPLE ACCUSED HIM OF FAKING THIS MAC MINI BATTERY TEST UNTIL HE PROVED IT DRAWS LESS POWER T...

PEOPLE ACCUSED HIM OF FAKING THIS MAC MINI BATTERY TEST UNTIL HE PROVED IT DRAWS LESS POWER THAN A DESK LAMP WHILE RUNNING LOCAL AI Most developers assume running local LLMs requires a noisy $3,000 GPU rig pulling 500 Watts straight from the wall. This creator

X.com07/27, 00:31Infrastructure

General AI value is flowing into Onchain AI China Open-Weight Labs → Inference providers → In...

General AI value is flowing into Onchain AI China Open-Weight Labs → Inference providers → Intelligent Routers Venice is eating more token share, DIEM secondary markets are getting established, and decentralized consumer inference are forming their structural

X.com07/26, 20:31Infrastructure

OPENAI SIGNED IT.

OPENAI SIGNED IT. GOOGLE SIGNED IT. NVIDIA SIGNED IT. ANTHROPIC DIDN’T. DAYS LATER, CHINA DROPPED KIMI K3 Coincidence? Maybe. But the timing couldn’t be more interesting. Kimi K3 is the first open 2.8T parameter MoE model with a 1M-token context window, native

X.com07/23, 22:03Infrastructure

Pretrained ViTs see the world in rich, dense detail.

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that enables transformer-based policies to consume dense tokens dir

X.com07/22, 23:46Infrastructure

FOUR INTEL ARC PRO B60 CARDS JUST TURNED A LOCAL AI PC INTO A 192GB VRAM POOL FOR MODELS ONE ...

FOUR INTEL ARC PRO B60 CARDS JUST TURNED A LOCAL AI PC INTO A 192GB VRAM POOL FOR MODELS ONE GPU CANNOT HOLD For the operator running recurring internal-code or document jobs, rented capacity becomes a scheduling problem before it becomes a model problem. Four

X.com07/22, 23:01Infrastructure

CHINESE DEV JUST STRAPPED A FULL GPU TO A $120 MINI-PC AND MADE $190K LAST YEAR WHILE HIS EX-...

CHINESE DEV JUST STRAPPED A FULL GPU TO A $120 MINI-PC AND MADE $190K LAST YEAR WHILE HIS EX-COFOUNDERS BURNED SERIES A ON ANTHROPIC BILLS Silver mini-PC on a cardboard base. GPU strapped to the side. UPS on top. Floral tablecloth edge visible. Build cost $550

X.com07/22, 23:01Infrastructure

A 2.8-TRILLION-PARAMETER MODEL JUST BEAT CLOSED AI AT THE ONE THING PEOPLE ACTUALLY PAY FOR: ...

A 2.8-TRILLION-PARAMETER MODEL JUST BEAT CLOSED AI AT THE ONE THING PEOPLE ACTUALLY PAY FOR: BUILDING. Kimi K3 reportedly took #1 in frontend coding, carries a 1-million-token context window, and can turn one prompt into games, 3D worlds, full products, or rep

X.com07/22, 22:47Infrastructure

In 1966, Stanford built Shakey — the world’s first AI robot.

In 1966, Stanford built Shakey — the world’s first AI robot. Modern AI was born right here. Almost nobody watches the original footage. No GPUs. No neural nets. No modern code. Just a 1-meter-tall box on wheels connected to room-sized mainframes. It was built

X.com07/21, 23:47Infrastructure

$56 versus $0.50.

$56 versus $0.50. same million tokens. one barrel costs 112x the other - and does maybe 20% more work. chamath said it plainly on cnbc: a "$50 barrel of intelligence" sitting next to a $1 one. here's how the premium quietly dies: step 1 → a rival ships 80-95%

X.com07/20, 21:02Infrastructure

THE PEOPLE WINNING WITH AI AREN'T SPENDING MORE.

THE PEOPLE WINNING WITH AI AREN'T SPENDING MORE. THEY'RE SPENDING LESS. Everyone talks about the $200/month subscriptions. Almost nobody talks about the hundreds of millions of free AI tokens available right now. The smartest builders don't start by buying the

X.com07/20, 20:01Infrastructure

2 MODELS LEFT IN THE ROTATION, ONLY 1 STILL EARNS FULL PRICE an engineer built a routing setu...

2 MODELS LEFT IN THE ROTATION, ONLY 1 STILL EARNS FULL PRICE an engineer built a routing setup to cut AI costs routine coding scripts go to a cheap model, under $1 per million tokens the hardest architecture problems still go straight to Fable 5, because nothi

X.com07/19, 20:47Infrastructure

A GOOGLE AI ENGINEER ACCIDENTALLY LEAKED HIS LOCAL OBSIDIAN KNOWLEDGE VAULT.

A GOOGLE AI ENGINEER ACCIDENTALLY LEAKED HIS LOCAL OBSIDIAN KNOWLEDGE VAULT. INSIDE - AN OFFLINE GRAPH EXPANDING ACROSS 18,000 INTERLINKED CLIENT FILES 18,000 knowledge nodes. 11,200 links. $0 cloud token bills opens the OpenClaw workspace. The screen displays

X.com07/19, 20:47Infrastructure

A 23-year-old Dev spent $4,699 once and cut $850 in monthly AI costs, and turned 1 desktop bo...

A 23-year-old Dev spent $4,699 once and cut $850 in monthly AI costs, and turned 1 desktop box into a business. Everyone talks about better AI models. Almost nobody talks about owning the computer that runs them. Cloud GPUs charge you every hour you think. A l

X.com07/02, 00:46Infrastructure

THIS DEVELOPER ADMITS HE'S GPU POOR WITH A 4-YEAR-OLD RTX 3090 - AND TESTS EVERY VRAM TIER SO...

THIS DEVELOPER ADMITS HE'S GPU POOR WITH A 4-YEAR-OLD RTX 3090 - AND TESTS EVERY VRAM TIER SO YOU KNOW EXACTLY WHAT YOUR HARDWARE CAN DO RTX 3060 with 12GB vs AMD BC250 with 16GB - one is a dedicated GPU with extra CPU and RAM, the other is an APU where one ch

X.com07/01, 23:16Infrastructure

temperature 0.5: the top token takes 96.97%.

temperature 0.5: the top token takes 96.97%. temperature 2.0: same token, 58.58%. The neighbors jump from 1.5% to 20%. The slider you drag without looking moves the distribution nonlinearly Broke it down on an 1889 Galton board:

X.com07/01, 21:33Infrastructure

A 31 YEAR OLD BUCHAREST OPERATOR JAMMED 8 USED NVIDIA T4s INTO A SUPERMICRO CHASSIS FOR $1,84...

A 31 YEAR OLD BUCHAREST OPERATOR JAMMED 8 USED NVIDIA T4s INTO A SUPERMICRO CHASSIS FOR $1,847, NOW SHIPS INFERENCE TO ROMANIAN SHOPIFY STORES FOR $11,240 A MONTH Basement workshop. Open Supermicro chassis on the bench. Eight slim T4 cards slotted like RAM sti

X.com07/01, 00:46Infrastructure

atomic[.]chat, a desktop app that runs LLMs locally, ran a very revealing comparison for Clau...

atomic[.]chat, a desktop app that runs LLMs locally, ran a very revealing comparison for Claude Sonnet 5, Claude Opus 4.8, Claude Sonnet 4.6, and GPT 5.5. Claude Sonnet 5 just matched GPT 5.5 on 3 physics coding demos at 6x lower cost. Also spent minimum numbe

X.com06/30, 21:16Infrastructure

A 29 YEAR OLD DENVER DEV PAID $42,000 FOR AN NVIDIA H100 RIG THAT SOLD FOR $400,000 NEW LAST ...

A 29 YEAR OLD DENVER DEV PAID $42,000 FOR AN NVIDIA H100 RIG THAT SOLD FOR $400,000 NEW LAST YEAR, NOW PULLS $43,200 A MONTH IN FINE TUNING REVENUE ethan is 29, denver basement office under a green ironing board, won an nvidia HGX H100 8 GPU server at a sunnyv

X.com06/30, 20:31Infrastructure

Inference will never be the same: Etched invented two new ways to massively improve compute c...

Inference will never be the same: Etched invented two new ways to massively improve compute cost, speed, and per watt efficiency: low voltage inference (more FLOPs) and cluster scale memory (memory/bandwidth) The combination runs trillion-parameter models at o

X.com06/30, 20:16Infrastructure

THIS CHINESE BUILDER TURNED 8 RETIRED TESLA P40s INTO A 192GB PRIVATE AI SERVER THAT CAN WORK...

THIS CHINESE BUILDER TURNED 8 RETIRED TESLA P40s INTO A 192GB PRIVATE AI SERVER THAT CAN WORK 5,760 GPU-HOURS EVERY MONTH. 00:06 he holds up one used Tesla P40 while seven more sit behind it, then installs the full stack inside a massive dual-Xeon server chass

X.com06/30, 19:16Infrastructure

PyTorch core engineer at Meta turned CUDA kernel writing into a sport in 13 minutes - better ...

PyTorch core engineer at Meta turned CUDA kernel writing into a sport in 13 minutes - better than $1500 GPU programming bootcamps. profile the kernel -> find the bottleneck -> rewrite -> benchmark -> merge the winning code into PyTorch. That loop is how the op

X.com06/30, 19:16Infrastructure

YOUR $599 MAC MINI BEATS A $3,800 GPU AT THE ONE THING THAT MATTERS Unified memory means the ...

YOUR $599 MAC MINI BEATS A $3,800 GPU AT THE ONE THING THAT MATTERS Unified memory means the model loads once and both processors read from the same pool A windows machine with double the price still copies data between ram and vram and chokes anyway - $130 te

X.com06/29, 21:31Infrastructure

Dan Fu co-wrote FlashAttention with Tri Dao.

Dan Fu co-wrote FlashAttention with Tri Dao. Then he co-built Hyena, Monarch Mixer, and ThunderKittens. Now he's distinguished researcher at Together AI. Four kernels you can find inside the inference stack running ChatGPT, Claude, and Gemini. One person, all

X.com06/29, 21:16Infrastructure

$3,999 OR $4,699.

$3,999 OR $4,699. SAME 128GB. SAME LOCAL INFERENCE. DIFFERENT LOGO ON THE BOX. AMD's Ryzen AI Halo developer platform. NVIDIA's DGX Spark. both run large models locally. both have 128GB of unified memory. same category, same use case, $700 apart. the video bel

OpenAI06/29, 08:17Infrastructure

OpenAI and Broadcom unveil LLM-optimized inference chip

OpenAI and Broadcom introduce Jalapeño, a custom AI chip built for LLM inference to improve performance, efficiency, and scale across AI systems.

X.com06/28, 22:46Infrastructure

COMPUTE AT THE EDGE IS NO LONGER A LUXURY Jensen Huang just dropped the Jetson Nano, and the ...

COMPUTE AT THE EDGE IS NO LONGER A LUXURY Jensen Huang just dropped the Jetson Nano, and the industry standard for AI inference efficiency has shifted. The unit economics: $249 entry cost. 25W power envelope. 70 trillion operations per second. This hardware is

X.com06/28, 21:46Infrastructure

CHINESE DEV BILLS $22K A MONTH RUNNING AI ON $80 CARDS BUILT FROM PS5 SILICON WHILE EVERY AI ...

CHINESE DEV BILLS $22K A MONTH RUNNING AI ON $80 CARDS BUILT FROM PS5 SILICON WHILE EVERY AI STARTUP SITS ON AN 8-MONTH H100 WAITLIST Lenovo office PC. AMD BC-250 sticking out the side. Same chip Sony ships in a PS5 now running inference for 3 paying clients.

X.com06/28, 21:46Infrastructure

A DEV TURNED A USED 2019 MAC PRO INTO A 35B LOCAL LLM SERVER BY PLUGGING IN ONE AMD EGPU AND ...

A DEV TURNED A USED 2019 MAC PRO INTO A 35B LOCAL LLM SERVER BY PLUGGING IN ONE AMD EGPU AND PUSHING THE RIG TO 52 TOKENS PER SECOND FOR UNDER $3,000 TOTAL he posts a video of the mac pro under the desk, one thunderbolt cable running to a Paladin eGPU enclosur

X.com06/28, 21:33Infrastructure

CLUSTERING 24 BC250 BOARDS FROM DEAD CRYPTO RACKS INTO A 384GB INFERENCE POOL HIT 1 TOKEN PER...

CLUSTERING 24 BC250 BOARDS FROM DEAD CRYPTO RACKS INTO A 384GB INFERENCE POOL HIT 1 TOKEN PER SECOND AND $5,000 ANNUAL POWER, THE TIER ZERO HACK ON YOUR MAP DOES NOT SCALE PAST 4 NODES 01:18 the operator looks at his bench, "it could run the model but you're t

X.com06/28, 20:45Infrastructure

THIS DEVELOPER RUNS A FULL LOCAL AI WITH WIKIPEDIA-SCALE KNOWLEDGE BASE - AND PAYS $3/MONTH W...

THIS DEVELOPER RUNS A FULL LOCAL AI WITH WIKIPEDIA-SCALE KNOWLEDGE BASE - AND PAYS $3/MONTH WHILE OTHERS PAY $300 open-frame AI workstation, LCD display showing real-time GPU load and memory usage, PCIe cards for local inference - and a monitor showing a knowl

X.com06/25, 22:35Infrastructure

So I almost added - Home GPU rig And it might be nice for self-sufficiency but again I still ...

So I almost added - Home GPU rig And it might be nice for self-sufficiency but again I still don't think local LLM stuff even comes close to cloud either in quality or performance (speed) or cost Making your home self-sufficient is nice though, I have: - 2x Te

X.com06/25, 22:15Infrastructure

We show how to get LLMs to communicate well in latent space instead of human language.

We show how to get LLMs to communicate well in latent space instead of human language. #ICML2026 spotlight (top 2%) Autoregressive latent thoughts -> KV transfer -> input-output alignment Speeds up inference by >4x by bypassing decoding and improves performanc

X.com06/25, 22:03Infrastructure

THIS CHINESE ENTREPRENEUR BUILT A WAREHOUSE-SCALE GPU FARM - AND NOW MAKES $4M/MONTH RENTING ...

THIS CHINESE ENTREPRENEUR BUILT A WAREHOUSE-SCALE GPU FARM - AND NOW MAKES $4M/MONTH RENTING TO AI STARTUPS 4,000 GPUs in warehouses the size of a Walmart - electricity contracts and cooling already in place - and an operator 6'6" tall for scale Ethereum merge

X.com06/24, 19:31Infrastructure

10,000 ASICS AT 100 DECIBELS BURN $1.6M/MO IN POWER.

10,000 ASICS AT 100 DECIBELS BURN $1.6M/MO IN POWER. THE SAME WAREHOUSE WITH 5,000 GPUS AT 85 DB CLEARS $4M FROM AI INFERENCE 100 decibels on the warehouse floor, workers wear industrial ear muffs all 8 hour shifts, you cannot hold a conversation 3 feet from o

X.com06/23, 19:45Infrastructure

$220/month in AI bills.

$220/month in AI bills. He paid it for two years without thinking. Then a client asked one question and walked out with $5,000 in 90 seconds. Watch at 0:21. That is where he points at the GPU sitting under the desk. One used RTX 3090. 24GB VRAM. Bought secondh

X.com06/18, 20:15Infrastructure

huggingface_hub v1.19.0 shipped Trusted Publishers: keyless CI/CD auth to the Hub via OIDC to...

huggingface_hub v1.19.0 shipped Trusted Publishers: keyless CI/CD auth to the Hub via OIDC token exchange. No more HF_TOKEN secrets sitting in your CI config. Here's how it works

X.com06/18, 06:30Infrastructure

Ondo is leading the tokenisation of stocks and ETFs the next big step will be robo advisors, ...

Ondo is leading the tokenisation of stocks and ETFs the next big step will be robo advisors, basically smart AI that is managing portfolios of stocks and ETFs. this is one area where actually Revolut was ahead of crypto and now crypto will have to catch up wit

X.com06/18, 05:45Infrastructure

SUPERCOMPUTERS IN 2026 ARE WAREHOUSES OF MAC MINIS.

SUPERCOMPUTERS IN 2026 ARE WAREHOUSES OF MAC MINIS. ONE ON YOUR DESK WITH KIMI API DELIVERS €300 RESEARCH REPORTS IN 15 MINUTES most people think running ai means a $10,000 gpu server or a cloud bill that grows every month. the developers building real busines

X.com06/15, 19:15Infrastructure

THIS IS WHAT LOCAL AI LOOKS LIKE WHEN IT STOPS BEING A TOY he is holding a 128GB AI mini PC b...

THIS IS WHAT LOCAL AI LOOKS LIKE WHEN IT STOPS BEING A TOY he is holding a 128GB AI mini PC built around AMD Strix Halo. that is the whole shift. for years, the answer to large models was simple: rent the cloud. pay the API. wait for the GPU. watch the meter.