Developer tools

Decode, build and debug faster: JWT, cron, encodings and more.

Reading in this category

Cursor Origin: what the GitHub rival actually ships

Cursor opened its Origin git forge to all paid plans on 17 August, the day GitHub broke. What it does, what it is missing, and who holds your code now.

Google's $10M Spirit Airlines Data Buy Slips to September 9

Google won a bankruptcy auction for Spirit Airlines' internal emails and Teams history at $10 million. The cabin crew union objected, so approval slipped to 9 September.

Cerebras CS-4: 750 PFLOPS from an overclocked wafer

Cerebras' CS-4 packs three WSE-3 Turbo wafers into one rack for 750 PFLOPS. The die is unchanged. Here's what the 30x claim really measures, and what's missing.

Groq's $350M reset: from Nvidia rival to Nvidia partner

Groq closed a $350M Series A on 17 August, reportedly at $3.5B, with Nvidia joining the round. The LPU challenger now sells Nvidia capacity. What that changes.

Stripe's $7B OpenRouter deal: what changes for your keys

Bloomberg says Stripe finalised a deal to buy OpenRouter for over $7B. Nothing is confirmed and nothing changed on your API keys. What Stripe is really buying.

Claude Code weekly limits drop a third after August 19

The +50% weekly limits promotion in Claude Code ends at 11:59 PM PT on 19 August 2026. That is a third off your headroom, not half. What changes, what does not.

Claude's text watermark: no opt-out, and no detector yet

Anthropic marks every Claude output with a SynthID style watermark. Global, no opt-out, weakest on code and short text, and the public detector has not shipped.

Nvidia's $105B OpenAI guarantee: a lease, not a loan

Nvidia disclosed residual value guaranties capped at $105 billion behind OpenAI's Ohio leases, next to $1.5 billion of actual equity. Here is what that obligates.

Nvidia's Q2 13F: $30B Intel, $21B SpaceX, zero buys

Nvidia's 13F for the June quarter shows a $63.44 billion equity book, up 3.4x. Every carried share count is unchanged, so the growth is marks plus one conversion.

Quantinuum Helios on Oracle Cloud: 98 qubits, $8M revenue

Quantinuum is installing a 98 qubit Helios inside an Oracle Cloud AI data center. No price, no availability date, and $8.0 million of quarterly revenue behind it.

Qwen3.8-27B ships Apache 2.0 with vision, the 2.4T doesn't

Alibaba's small open model got the permissive licence and a vision encoder. The 2.4T flagship weights are text only under a custom licence. What that changes.

Copilot's one app merge: what goes away on August 18

Microsoft folded its two Copilot apps into one. The merge is automatic. Group Chat, Podcasts and consumer Deep Research retire on 18 August.

Palmyra X6 runs on GLM-5.2, priced $2 in and $8 out

Writer's new flagship is post-trained GLM-5.2 from Z.ai. It scores 0.87 against Opus 4.8's 0.86, and that number comes from Writer's own nine evaluations.

GLM-5.3: open weights in two weeks, breaking API change now

Z.ai shipped GLM-5.3 on 14 August. The weights land in about two weeks, and thinking can no longer be disabled, so swapping the model ID breaks non-thinking calls.

GPT-5.6 Sol Ultrafast: 14x on the chart, 5.6x end to end

OpenAI previewed Ultrafast on 13 August, a Cerebras backed API tier running GPT-5.6 Sol at up to 750 tokens a second. End to end it is 5.6x, and no price is published.

Uber and Pony.ai: 2,000 robotaxis, four cities unnamed

Pony.ai and Uber announced more than 2,000 robotaxis across five European cities on 14 August 2026. No city names, no dates, and the fleet it grows from is 10 cars.

Grok 4.6: same $2/$6, and a 200K pricing cliff

xAI shipped Grok 4.6 on August 12 at the same $2/$6 as 4.5, plus five index points. Two costs moved quietly: the cache rate, and a 200K repricing cliff.

Twitch's AI training toggle is on, and it covers your chat

Twitch switched on Amazon genAI training by default on 12 August. The toggle hides in Security and Privacy, and your setting decides for everyone in your chat.

Gemini 3.7 Flash: half price, and it doubles on 1 January

Google shipped Gemini 3.7 Flash on 13 August at $0.75/$3.75 per million tokens. The pricing page says that rate runs out on 31 December 2026.

Nemotron 3.5 Lightning: 3B active, 350 tokens a second

NVIDIA's 30B mixture-of-experts model runs on 3B active parameters, and the post-training data shipped too. The speed survives outside NVIDIA. The context claim does not.

DeepSeek V4 Pro 0813 is live, at 3x the Flash price

DeepSeek pointed deepseek-v4-pro at the 0813 build on 12 August. Same endpoint, same price, no change log entry, and no 0813 weights on Hugging Face.

Intel's $20B share sale: 14A still has no named customer

Intel priced a $20 billion stock sale on 10 August, its first since 1971. The filing names no 14A foundry customer, and that is the number that actually matters.

Anthropic's Theseus: a data center deal with no numbers

Anthropic, Macquarie and GIC launched Theseus Infrastructure on 10 August. No capital figure, no megawatts, no sites. The structure is the actual news.

NVIDIA's $500B compute financing and the 25% backstop

NVIDIA signed MOUs with six asset managers on 10 August to mobilize over $500 billion for AI compute. It may also cover 25% of the residual value risk.

Source Foundry: $400M and a $5B bet against ASML

Situational Awareness put $400 million into Source Foundry at a $5 billion valuation. The startup wants to build EUV lithography tools. Here is what that money buys.

Claude Code auto mode is the default from August 14

Auto mode becomes the default permission mode in Claude Code on 14 August for Pro, Max and Team. What the classifier blocks, what it misses, and who still pays.

Meta Muse Glimmer: the 30B agent needs 24GB of VRAM

Meta released Muse Glimmer on 10 August: 30B dense, Apache 2.0, open weights. It does run locally, but the floor is a 24GB card and the headline speed is not yours.

Genesis-Science-1: DOE Opened a Portal, Not a Model

The DOE launched its Genesis Open Models Initiative on 7 August, first model Genesis-Science-1. No size, no license, no weights yet. One real deadline.

Wayve and Uber cleared in London: 15 cars, driver inside

TfL granted private hire licences to Wayve self-driving Ford Mustang Mach-E cars on 5 August 2026. Up to 15 of them, one year each, human driver on board.

ByteDance's 10T model is pretraining, not a product

The Financial Times says ByteDance is pretraining a model of up to 10 trillion parameters on about 30,000 GPUs. No name, no active count, no date, no confirmation.

AMD buys Taalas: the HC1 etches Llama 3.1 into silicon

AMD is acquiring Taalas, whose HC1 chip burns Llama 3.1 8B weights into TSMC 6nm mask layers instead of HBM. What that buys, and what it costs you.

Terafab: what Tesla and SpaceX's $16.8B first phase buys

Tesla and SpaceX confirmed Terafab in Grimes County, Texas. $16.8 billion for phase one, 3,000 jobs, no process node and no date. What is real and what is a slide.

IonQ Evergreen-05: DARPA funds 25 clocks, not 125

IonQ's updated DARPA release promises 125 Evergreen-05 atomic clocks. The signed $28M covers 25 of them. We read the structure, and the specs.

GPT-5.6 Sol in ChatGPT is not the Sol in your API calls

OpenAI retuned GPT-5.6 Sol for ChatGPT chat only on 6 August. Work, Codex and your API calls keep the old one. Free users move to Luna with unlimited chats.

Anthropic's custom silicon team: no chip, no timeline

Anthropic confirmed an in-house silicon team on 5 August 2026. The roles are public, first silicon is future tense, and AWS, Google, Nvidia and AMD all stay.

Meta Muse Code: Muse Spark 1.2 and the 21x cheaper tier

Meta shipped Muse Code in beta on 5 August, a terminal coding agent on Muse Spark 1.2. It tops none of Meta's own charts. The pitch is a 21x cheaper tier.

Jeff Dean leaves Google: what changes for Gemini devs

Jeff Dean, Sanjay Ghemawat, Oriol Vinyals and Quoc Le are off to Discovery Loop, and Hassabis steps back. What it does to your Gemini integration: nothing yet.

Unitree IPO prices at $9B: 73.6% of humanoids go to labs

Unitree priced its Shanghai IPO at 150.80 yuan, valuing it near $9 billion. The prospectus says 73.6 percent of its humanoid revenue came from research and education.

Shieldstral 1.0: the 3B guard model that ties a 20B

Mistral shipped Shieldstral 1.0 on 4 August, a 3B open-weights moderation model for text and images. It ties a 20B on text F1 and leads on multimodal.

Grok Voice Think Fast 2.0: grok-voice-latest costs 60% more

xAI moved the grok-voice-latest alias to Think Fast 2.0 today. Same connection string, $0.05 to $0.08 a minute. What you get for it, and whether the speed pays.

Google Assistant shuts down September 4: what you lose

Google starts removing Assistant from phones, watches, headphones and Android Auto on 4 September 2026. There is no switching back, and Gemini has gaps.

Qwen 3.8 Max: is it open source? Not yet

Headlines call Qwen3.8-Max open source. Today it is API only, the weights are promised next week, and no license has been named. What that really means.

MiniMax H3 open weights: not licensed in the EU, UK or US

MiniMax put the 33B H3 video weights on Hugging Face on August 3. The license names the EU, the UK, South Korea and the US as excluded territories.

Bending Spoons buys Airtable for $1.285B: what changes

Bending Spoons is buying Airtable for $1.285 billion, all cash. The deal is confirmed. What happens to your bill is not, so here is the track record.

MacBook Air M5 shortage: $1,299 and a six week wait

Apple raised the MacBook Air to $1,299 in June and still cannot keep it in stock. US delivery runs two to six weeks, and the higher RAM configs wait longest.

Waymo's Santa Monica depot ban: 11pm to 6am until trial

A judge barred Waymo from running its two Broadway charging lots between 11pm and 6am. The robotaxi bottleneck turns out to be the depot, not the driving.

EU AI Act Article 50 is live: what you must label

The AI Act transparency rules applied on 2 August 2026. The delay everyone read about covered different articles. What actually binds you now, and what still has no test.

Qualcomm's double-digit chip price rise starts September 1

Qualcomm confirmed a double-digit price rise on chips shipping after September 1. It never named a percentage, and it is small next to the memory bill.

Inkling-Small: 276B that beats the 975B at coding

Thinking Machines shipped Inkling-Small on 30 July: 276B total, 12B active, Apache 2.0. It out-codes the 975B flagship, and forgets half of what that one knew.

OpenAI Astra: what the ten Lean proofs actually prove

OpenAI's unreleased Astra produced ten results on decade-old maths problems for about $2,000. The Lean certificates are public and they compile. What that settles.

DeepSeek V4-Flash-0731: same API, new model

DeepSeek swapped the weights behind deepseek-v4-flash on 31 July. Same endpoint, same $0.14 price, Terminal Bench 2.1 up from 61.8 to 82.7.

Google pulled Nano Banana 2 from Earth after one day

Google shipped AI image generation into Google Earth on 30 July and rolled it back on 31 July. The SynthID watermark only covers the pixels it generated.

Zoox wins Part 555: paid robotaxi, no steering wheel

NHTSA granted Zoox a two-year Part 555 exemption on 31 July, up to 2,500 vehicles a year. First time a passenger vehicle with no driver controls can charge fares.

Gemini Robotics 2: what Google's own benchmarks say

Google DeepMind shipped Gemini Robotics 2 on 30 July with a demo reel and a score table. The table says a dustpan works 32 percent of the time. We read both.

Amazon's $1.8M Claude bill: 860% over, 5 months unseen

The FT reported a failed Amazon tool that burned $1.8 million in Claude Sonnet tokens. At list price that is 120 to 600 billion tokens, unnoticed.

FCC robot ban: the 4.4 lb rule that catches vacuums

The FCC put foreign-produced robots on its Covered List on 28 July. The definition starts at 4.4 lb with one sensor and 200 kbps, so it catches robot vacuums.

GPT-5.6 price cut: Luna drops 80%, Sol stays at $5

OpenAI cut GPT-5.6 API prices on 30 July. Luna drops 80% to $0.20 / $1.20 and Terra 20%, while Sol does not move. What it does to your bill.

MCP 2026-07-28: what the stateless core removes

MCP 2026-07-28 makes the protocol stateless and removes sessions outright. The twelve month deprecation policy in the same release does not cover that list.

Nvidia SSI deal: $5B reported, 10x compute on Vera Rubin

Nvidia and Safe Superintelligence announced a partnership on 27 July. What the release actually says, what the $5 billion rests on, and what 10x leaves out.

OpenAI's free GPT-5.6 for researchers: what you get

OpenAI opened applications on 29 July for free frontier model access at universities. The real terms: five seats, twelve months, ChatGPT Pro rate limits.

Pixel 11 price increase: the RAM crisis, in numbers

Google confirmed Pixel 11 prices are going up and blamed a supplier-driven RAM crisis. The two numbers it cited divide to 4.3x, not the sixfold everyone printed.

Laguna S 2.1: 118B open weights on one DGX Spark

Poolside ships Laguna S 2.1, a 118B MoE with 8B active and a 1M context, under an open licence. The weights fit one DGX Spark at four bits. The 1M window does not.

Tesla Optimus V3: production starts, reveal still missing

Tesla is installing Optimus production lines at Fremont for a robot it has never shown. What the Q2 call actually confirmed, and why every V3 spec sheet is fiction.

NVIDIA and SK's $500B deal: what the 2GW actually commits

SK Group and NVIDIA announced a $500 billion-plus partnership on July 24. The instrument is letters of intent. What is named is a 2GW AI factory for 2027, plus HBM4.

IBM buys HRL: 18 spin qubits and a 300mm fab

IBM is buying HRL Laboratories from Boeing and GM. The release has no price and no qubit count. We read the April paper that does: 54 dots, 18 qubits.

iOS 26.6 is out: the Spotlight index it quietly builds

Apple shipped iOS 26.6 on July 27. One line of the release note matters: it rebuilds the Spotlight index that the new Siri in iOS 27 will read.

Kimi K3 open weights land July 27: the 1.4TB catch

Moonshot ships the Kimi K3 weights by July 27. At 2.8 trillion parameters in MXFP4 that is about 1.4 TB, and Moonshot itself recommends 64 or more accelerators.

Cursor agent swarm: SQLite for $1,339 and the 15x catch

Cursor swarmed agents at the 835-page SQLite manual and got a working Rust database for $1,339. The 15x cost gap being quoted needs a run Cursor left out.

AMD Helios vs Vera Rubin NVL72: the MI455X numbers

AMD launched Helios on July 23: 72 Instinct MI455X GPUs, 31TB of HBM4, and a claim of 30% more tokens per dollar than Vera Rubin NVL72. We read the footnotes.

Open Weights letter: 25 signers, and what it changes

On July 24, 25 companies signed Open Weights and American AI Leadership. Who signed, who pointedly did not, and what it changes if you ship on downloadable models.

Etched raises $300M at $10.3B: where did Sohu go?

Etched raised $300M at a $10.3B valuation on July 23. Its own site no longer says transformer, or Sohu. What it is really shipping, and what stays unproven.

FLUX 3 is here: video, audio, open weights later

Black Forest Labs launched FLUX 3 on July 23: one model for video with audio, image and robot actions. Early access only, and the open weights land later in 2026.

Claude Opus 5 is here: same price, thinking on by default

Anthropic shipped Claude Opus 5 on July 24 at the same $5/$25 price as Opus 4.8, pitched as Fable 5 intelligence at half the cost, with thinking on by default.

PsiQuantum's $125M DARPA deal: what it validates

PsiQuantum signed a $125M DARPA agreement on July 22. It funds verification of a quantum pathway, not a working machine. We read what it actually buys.

OpenAI Presence: the enterprise AI agent platform

OpenAI launched Presence on July 22, a managed platform for governed voice and chat agents. What it is, who it is for, and what the 75% figure really means.

Anthropic's $1.5B author settlement: what it changes

A judge gave final approval to Anthropic's $1.5B author settlement on July 20. What it actually settles, the fair-use question it leaves open, and what changes for you.

ChatGPT ads are live: OpenAI's self-serve Ads Manager

OpenAI opened a self-serve Ads Manager for ChatGPT on July 22. Who sees the ads, where they sit, the budget confusion, and what it actually means for you.

Claude Cowork Record a Skill: teach agents by demo

Anthropic shipped Record a Skill in Claude Cowork on July 21. What it does, which plans get it, and the screen-capture catch nobody is flagging.

NVIDIA Vera Rubin: what BMS's first AI factory really is

Bristol Myers Squibb is deploying an NVIDIA DGX SuperPOD on eight Vera Rubin NVL72 systems. It's a build commitment, not a running cluster. What Rubin really is.

Qwen-Image-3.0 is here: no weights, no benchmarks

Alibaba's Qwen-Image-3.0 went live July 21 claiming 10px text and live-data infographics, but with no weights, no benchmark and no model card. The honest read.

Gemini 3.6 Flash: cheaper than 3.5, better on paper

Google shipped Gemini 3.6 Flash on July 21 at $1.50/$7.50 per million tokens, cheaper than 3.5 Flash and higher on its own benchmarks. The Pro is still in testing.

Robostral Navigate: Mistral robots on one camera

Mistral's first robotics model steers robots from one RGB camera and plain-English commands. Here's what the 8B Robostral Navigate really does, and where it stops.

Did Claude Fable 5 disprove the Jacobian Conjecture?

An Anthropic number theorist used Claude Fable 5 to post a public counterexample to the 85-year-old Jacobian Conjecture. What is real, and what still needs checking.

Claude Code drops Node for Bun 1.4.0 in Rust

Since v2.1.181, Claude Code ships as a native binary running Bun, ported from Zig to Rust with Claude. What changed for you, the real numbers, and the slop row.

Humanoid HMND 01: what the robot unicorn really ships

London's Humanoid just hit a $1.2B valuation. Here's what its HMND 01 Alpha robot actually does, the specs that are real, and what's still just a claim.

Qwen3.8 preview is live: the 2.4T open weights are not

Alibaba's Qwen3.8-Max-Preview went up on July 19 with a 2.4T-parameter claim and open weights promised soon. Here's what you can run today, and what's still a tweet.

Fireworks AI $1.5B Series D: what it means for open models

Fireworks closed a $1.505B Series D at a $17.5B valuation, Nvidia among the backers. A raise ships no code. Here is what it signals for teams serving open weights.

Gemini 3.5 Pro missed July 17: launch vs leak

Google's Gemini 3.5 Pro slipped its reported July 17 target, a third miss. The 2M context, Deep Think and pricing are all leaks. Here is what Google actually said.

Claude Fable 5 goes permanent July 20: the fine print

Anthropic makes Claude Fable 5 permanent from July 20, but only Max and Team Premium include it, at 50% of limits. Pro and Team Standard pay per token.

NVIDIA Ising Decoder ColorCode 1: what 347x really means

NVIDIA's open color code pre-decoder claims 347.7x better logical error rates. We read the chart, and the headline point at d=31 is extrapolated, not measured.

Apple Intelligence China: approved, not shipped

China cleared Apple Intelligence on July 15 with Alibaba Qwen inside. No launch date came with it. What the filing actually covers, and what it does not change.

NVIDIA Jetson T3000 and T2000: Thor for mainstream robots

NVIDIA's new Jetson T3000 and T2000 bring Thor-class AI to smaller, cheaper robots. Here's the real spec gap, the Q1 2027 catch, and what's still hype.

Thinking Machines Inkling: what the 975B open model is for

Thinking Machines released Inkling, a 975B open-weights MoE with 41B active params. It loses plenty of benchmarks on purpose. Here is what it is actually built for.

iOS 27 public beta: the new Siri AI is finally here

Apple opened the iOS 27 public beta with a rebuilt Siri that holds a conversation and acts in your apps. Here is who can actually use it, and the real catches.

Kimi K3 is live: the confirmed specs and the leak noise

Moonshot's Kimi K3 hit the Kimi app and API on July 16 with a 1M context. The 2.8T params, benchmarks and prices are still leak-grade. What to trust today.

Apple v. OpenAI: what the trade secrets suit changes

Apple sued OpenAI on July 10 over alleged trade secret theft. What the complaint actually claims, what stays unproven, and what it changes for your work.

Nvidia halves its Asia AI chip buyer list: who's cut

Nvidia cut more than half its authorised AI chip buyers in Singapore, Malaysia and Japan. What the white list check involves, who failed it, and if it reaches you.

DeepSeek V4 is here: deepseek-chat breaks July 24

DeepSeek V4 is graduating from preview, and the deepseek-chat and deepseek-reasoner API aliases stop working after July 24, 2026. What to change, and the one trap.

Seedream 5.0 Pro: the image model that ships you layers

Seedream 5.0 Pro (ByteDance, July 8) splits an image into 10+ editable transparent layers and renders in-image text in ten-plus languages. Where it wins, and slips.

Meta Iris: what its new AI chip changes for you

Meta's new AI chip, Iris, enters production in September, built with Broadcom on TSMC. Here's what it really changes for you, and what's still just hype.

Meta Muse Spark 1.1: cheap agentic coding, real caveats

Meta shipped Muse Spark 1.1 on July 9, its first paid model API: a 1M-token agentic coder at $1.25/$4.25. What it wins, where it trails, and the catch.

TypeScript 7.0 is here: the Go rewrite is 10x faster

TypeScript 7.0 shipped July 8, the compiler rewritten in Go and about 10x faster on full builds. What really changed, and the one catch for Astro and Svelte users.

GPT-5.6 goes public July 9: Sol, Terra or Luna

OpenAI is releasing GPT-5.6 to everyone from July 9 after a US safety review. Sol, Terra and Luna tiers, the prices, and which one to actually use.

Tencent Hy3: what a free 295B open model really gives you

Tencent's Hy3 is an open 295B MoE under Apache 2.0, free on OpenRouter until July 21. Here's what it really is and the hardware you need to run it.

Claude Fable 5 vs Grok 4.5: is 8x cheaper worth it?

Fable 5 wins the hard coding benchmarks; Grok 4.5 costs about 8x less on output and burns far fewer tokens. The gap, the price, and which one to actually run.

xAI Grok 4.5 is here: cheap, fast, coding-first

xAI shipped Grok 4.5 on July 8 at $2/$6 per million tokens, the cheapest serious coding model. What it wins, where it trails Fable 5, and the EU catch.

OpenAI GPT-Realtime-2.1: what changed for voice agents

OpenAI shipped gpt-realtime-2.1 and a mini reasoning model on the Realtime API. What actually changed, the 25% latency cut, and when to pick each.

OpenAI Jalapeño: what its first chip changes for you

OpenAI's first custom chip, Jalapeño, is an inference accelerator built with Broadcom. Here's what it really changes for you, and what's still just hype.

Claude Fable 5 is back: what each effort level costs

Claude Fable 5 is back worldwide. Its rate is fixed at $10/$50, but effort (low to max) multiplies real cost by 7x. The numbers per level, and when each is worth it.

Claude Sonnet 5 vs Fable 5: is the top model worth 5x?

Fable 5 costs five times Sonnet 5 right now. It buys a 17-point SWE-bench Pro lead and long-horizon depth. Where the top model earns it, and where Sonnet 5 is enough.

Claude Sonnet 5 vs Opus 4.8: cheaper, nearly as good

Claude Sonnet 5 costs 40 to 60 percent less than Opus 4.8 and closes most of the gap, even winning agentic coding. Where Opus still wins, and which to pick.

Qwen 3.7 Max vs GLM-5.2: benchmarks, price and the catch

Qwen 3.7 Max vs GLM-5.2, fact-checked: they tie on coding benchmarks, but GLM is cheaper and open-weight while Qwen Max wins math, reasoning and long autonomous runs.

Qwen 3.7 local: what you can actually run offline

Qwen 3.7 Max is API-only, so you can't run it locally yet. Here's how to run Qwen offline today with Ollama and the open Qwen models, sized to your hardware.

GPT-5.6 vs GPT-5.5: what actually changed

GPT-5.6 is not one model but a new Sol, Terra and Luna tier system. What changed from GPT-5.5: prices, the coding SOTA, the ultra subagent mode, and whether to switch.

GLM-5.2 vs GPT-5.5 and Opus 4.8: benchmarks and price

GLM-5.2 vs GPT-5.5 and Claude Opus 4.8, fact-checked: it beats GPT-5.5 on coding and costs about 6x less than Opus, but Opus still leads the hardest tasks.

How to search text in files with grep

Search text in files with grep: grep -r for a whole tree, -i to ignore case, -n for line numbers, -l for filenames. The flags that matter, with a real example.

How to download a file and test an API with curl

Use curl to download files (-O, -L) and test APIs (-i, -X, -d) from the command line. The flags that matter, a POST with a JSON body, and how to read the response.