explore
Pick a thread.
Every story we have published, by topic. Tap a tile to filter, with no reloads, just the thread you want to pull.
All stories
120 stories
DeepSeek is retiring its flagship into its cheap model, and the cheap one has vision
DeepSeek released V4.1-Flash on 10 September 2026 and said it will route every V4 Pro request to it four days later, at the cheaper price. The pricing table carries the odd part: the model being retired does not support vision, and the one replacing it does.
Ahmad J · Sep 10, 2026 · 5 min read

Google's cheapest AI plan gets voice in Gmail and Keep, not in Docs
Google added four features to its consumer AI plans on 9 September 2026, and the conditions in brackets are the story. Three of the four skip the entry paid tier entirely, the fourth reaches it in two apps out of three, and on Google's own plans page every tier runs the same model.
Ahmad J · Sep 9, 2026 · 6 min read

AWS's new G7 instances won its own benchmark with half the GPUs
AWS benchmarked its Blackwell-based G7 instances against G5, G6 and G6e on 30B mixture-of-experts models. The two-GPU box beat the four-GPU boxes, the cheapest configuration and the fastest one turned out to be different machines, and the same model cost 2.5 times more per token on retrieval traffic than on chat.
Ahmad J · Sep 8, 2026 · 6 min read

AWS Bedrock's zero data retention has two exceptions
Bedrock stores nothing by default. Two Anthropic models will not run unless you switch that off, and the switch is an API call with no console for it.
Ahmad J · Sep 8, 2026 · 7 min read

Mistral raised €3 billion on open weights. Read the licences before you ship.
Mistral's Series D values it above €21 billion and its models really are downloadable. But four different licences sit behind the words open weight, and the frontier one stops applying once your company passes $20 million of revenue in a month.
Ahmad J · Sep 8, 2026 · 7 min read
More stories
Article · softwareWindows 10 has ended. What your options actually areSep 8, 2026 · 6 min read
Article · hardwareWhy your computer is slow, and what actually fixes itSep 7, 2026 · 8 min read
News · businessGoogle's Lyria 3.5 has one price, and the 30 second clip stayed on Lyria 3Sep 6, 2026 · 6 min read
Guide · aiWhich local LLMs fit in 8, 12, 16, 24 or 32GB of VRAMSep 6, 2026 · 8 min read
Guide · softwareHow to fix a slow computer on Windows 11, rankedSep 5, 2026 · 8 min read
News · securityOpenAI's cyber model has one price and no way to lower itSep 5, 2026 · 8 min read
News · businessGPT-6 Astra has a million-token window and a price cliff at 272,000Sep 5, 2026 · 6 min read
Guide · aiHow to Choose a Laptop for Local AISep 5, 2026 · 4 min read
Tutorial · softwareEvaluate whether a model is good enough for your taskSep 4, 2026 · 3 min read
Guide · aiChoosing the right model size for your taskSep 4, 2026 · 4 min read
News · businessGemini Flash is half price until 31 December, and Google published the date it endsSep 3, 2026 · 6 min read
News · softwareAsked which software to buy, an AI assistant cited a demo vendor's blog more often than GartnerSep 3, 2026 · 8 min read
Article · hardwareOn-Device AI on Phones: Privacy and Latency, Not HypeSep 3, 2026 · 5 min read
News · aiAI Agents Are Moving From Demos to Narrow JobsSep 3, 2026 · 3 min read
News · securityMistral trains on your chats by default. Turning it off takes three separate switchesSep 2, 2026 · 6 min read
News · scienceMicrosoft distilled a billion-parameter pathology model into 22M parametersSep 2, 2026 · 4 min read
News · aiGemini can now skim a video instead of watching every frameSep 2, 2026 · 4 min read
News · aiWorld Labs' Atlas wins six of seven 3D benchmarks, and you cannot run itSep 2, 2026 · 5 min read
News · aiClaude Fable 5.1 costs 25% less, and its per-token price did not moveSep 2, 2026 · 5 min read
News · aiSmall Models Are Quietly Taking Over the Easy WorkSep 2, 2026 · 3 min read
Guide · aiWhich AI subscription should you actually pay for?Sep 2, 2026 · 7 min read
News · aiDeepSeek's first V4 vision model is MIT-licensed, and 307 GBSep 1, 2026 · 5 min read
News · aiA cheap model solved open math problems, once Google put a team around itSep 1, 2026 · 5 min read
Review · hardwareUSB AI Accelerators: The External Stick for Running ModelsSep 1, 2026 · 3 min read
Guide · hardwareHow to Choose Hardware for Local AISep 1, 2026 · 3 min read
Article · hardwareWhy Future-Proofing a Computer Is Mostly a MythAug 31, 2026 · 3 min read
Guide · smartphonesHow to Choose a Phone That LastsAug 31, 2026 · 5 min read
Guide · securityIs your router still getting security updates?Aug 31, 2026 · 13 min read
Guide · hardwareWhen does your laptop stop getting security updates?Aug 31, 2026 · 11 min read
Article · businessOpen-Source vs Closed AI Models: The Business Risk Your Team Isn't Pricing InAug 17, 2026 · 4 min read
Article · hardwareComputational Photography: How Phone Cameras Use AIAug 17, 2026 · 5 min read
Article · smartphonesApple Already Reads Your Texts With AI. Let Developers Read One Line of Them.Aug 17, 2026 · 4 min read
Article · hardwareWhat Thermal Throttling Is and Why Thin Devices Slow DownAug 17, 2026 · 5 min read
Article · smartphonesWhat actually happens when your phone stops getting security updates?Aug 16, 2026 · 11 min read
Guide · smartphonesHow long does Apple support iPhones?Aug 16, 2026 · 11 min read
Guide · smartphonesHow long will my phone get security updates?Aug 16, 2026 · 17 min read
Article · aiSmall AI Models Are Quietly Winning in ProductionAug 14, 2026 · 4 min read
Article · aiAI agents in production: the honest 2026 state of playAug 13, 2026 · 8 min read
Guide · hardwareHow to Build an AI Workstation on a Tight BudgetAug 13, 2026 · 3 min read
News · chipsThe NPU quietly became standard hardwareAug 13, 2026 · 3 min read
Tutorial · softwareHow to run a local LLM on your own machine with OllamaAug 13, 2026 · 4 min read
Article · aiLocal vs. Cloud AI Processing: The Real Trade-OffsAug 12, 2026 · 4 min read
Article · hardwareWhat Unified Memory Actually Changes for a LaptopAug 12, 2026 · 4 min read
News · aiOpen-Weight Models Changed Who Controls the AI StackAug 12, 2026 · 3 min read
Tutorial · aiHow to fine-tune a small language model with LoRAAug 12, 2026 · 4 min read
Tutorial · softwareHow to structure prompts for reliable, parseable LLM outputAug 11, 2026 · 4 min read
Tutorial · softwareSet up GPU drivers and the toolkit for local AI workAug 11, 2026 · 3 min read
Article · aiMemory Bandwidth, Not Compute, Limits Local LLM SpeedAug 11, 2026 · 5 min read
Article · hardwareThe US–China Chip Export Controls: Where They Restrict, and Where They Don'tAug 11, 2026 · 5 min read
Tutorial · softwareServe a local model as an API endpointAug 11, 2026 · 3 min read
Article · aiWhat an AI Benchmark Actually Measures (And Why Leaderboards Mislead You)Aug 10, 2026 · 5 min read
Article · aiThe Real Cost of AI Is Inference, Not TrainingAug 10, 2026 · 9 min read
Article · scienceHow AI Cracks Hard Problems in Drug Discovery, Materials Science, and Climate ModelingAug 10, 2026 · 5 min read
Tutorial · aiHow to build a basic RAG pipeline for a local LLMAug 10, 2026 · 4 min read
Article · securityThe Local Illusion: The Real Security Risks of Running a Local LLMAug 9, 2026 · 6 min read
Article · aiHow LLM Context Windows Actually Work (and Why Bigger Isn't Always Better)Aug 9, 2026 · 4 min read
Article · scienceScience at Scale: How AI Is Restructuring Which Questions Researchers Can Afford to AskAug 9, 2026 · 10 min read
Article · businessWhat a Model Actually Costs to Run in Production: A Back-of-Envelope Framework for TeamsAug 9, 2026 · 8 min read
Article · aiPrompt Caching: How It Actually Cuts Your LLM API BillAug 8, 2026 · 3 min read
Article · aiWhy GPUs Beat CPUs for AI InferenceAug 8, 2026 · 7 min read
Article · roboticsThe Real Robotics Stack: Where Sensors, Compute, and Middleware Actually BreakAug 8, 2026 · 11 min read
Article · aiLoRA and QLoRA Fine-Tuning: How to Customize LLMs Without Burning Your BudgetAug 8, 2026 · 8 min read
Article · aiTokens, Explained: How Language Models Read Your Text and How You're BilledAug 7, 2026 · 8 min read
Article · hardwarePCIe 4.0 vs 5.0 vs Thunderbolt for AI Workloads: Where the Generational Upgrade Actually MattersAug 7, 2026 · 9 min read
Article · aiHow AI Coding Assistants Actually Work Under the HoodAug 7, 2026 · 6 min read
Article · softwareVector Databases: How They Actually Work, and When You Don't Need OneAug 7, 2026 · 4 min read
Article · aiQuantization Explained: What Q4, Q8, and FP16 Actually Do to a Local ModelAug 6, 2026 · 8 min read
Article · hardwareCUDA Lock-In Is Real: A Precise Cost Accounting of What Switching GPU Vendors Actually BreaksAug 6, 2026 · 9 min read
Article · roboticsHow Robots Are Really Trained: The Sim-to-Real Gap Is Not a Bug You Can PatchAug 6, 2026 · 10 min read
Article · aiMixture-of-Experts Models: How They Work and Why They Cut Inference CostsAug 6, 2026 · 10 min read
Article · aiThe Real Difference Between MCP, Function Calling, and Agent LoopsAug 5, 2026 · 3 min read
Article · aiHow Much VRAM You Actually Need to Run a Local LLMAug 5, 2026 · 7 min read
Article · securityThe Real Privacy Audit: What Data Your AI Coding Assistant Sends HomeAug 5, 2026 · 11 min read
Article · aiHow to Evaluate a Local LLM for a Real Task: A Repeatable Testing FrameworkAug 5, 2026 · 11 min read
Guide · aiThe best local LLM runners in 2026: Ollama, LM Studio, vLLM, and moreAug 4, 2026 · 9 min read
Article · securityA door closes, a window opens: Project Zero's 0-click chain reaches the Pixel 10 kernelAug 4, 2026 · 5 min read
Guide · smartphonesThe best smartphones for on-device AI in 2026Aug 3, 2026 · 11 min read
Guide · aiThe best home-server hardware for self-hosting AI in 2026Aug 3, 2026 · 13 min read
Guide · aiThe best cloud GPU providers for AI training in 2026Aug 3, 2026 · 11 min read
Guide · aiThe best AI note-taking and writing tools in 2026Aug 2, 2026 · 11 min read
Guide · aiThe best Macs for local AI and machine learning in 2026Aug 2, 2026 · 11 min read
Guide · aiThe best cloud hosting for running AI models in 2026Aug 2, 2026 · 11 min read
Guide · aiThe best vector databases for RAG in 2026Aug 1, 2026 · 12 min read
Guide · aiThe best CPUs for AI development workstations in 2026Aug 1, 2026 · 13 min read
Guide · aiThe best AI image generators in 2026Aug 1, 2026 · 13 min read
Article · chipsWhy only a few foundries make the leading-edge chipsJul 31, 2026 · 8 min read
Article · chipsWhat a Nanometer Process Node Really MeansJul 31, 2026 · 9 min read
Article · chipsCoWoS, SoIC, and Foveros: How Advanced Chip Packaging Actually WorksJul 30, 2026 · 11 min read
Article · chipsTSMC's Pricing Power: How One Foundry Sets the Cost of Every Advanced ChipJul 30, 2026 · 10 min read
Article · chipsHow EUV Lithography Works, in Plain TermsJul 30, 2026 · 8 min read
Article · chipsAI Inference on the Edge: How Embedded Chips in Cars, Cameras, and Appliances Actually WorkJul 29, 2026 · 12 min read
Article · chipsRISC-V vs ARM vs x86: The Honest ComparisonJul 29, 2026 · 10 min read
Article · chipsHBM Explained: Why High-Bandwidth Memory Is the Real Bottleneck in AI ChipsJul 29, 2026 · 12 min read
Article · hardwareWhy Battery Life Is a Chip and Software StoryJul 28, 2026 · 11 min read
Article · aiModel Distillation: How Small Models Learn to Punch Above Their WeightJul 28, 2026 · 9 min read
Article · chipsWhat Chiplets Are and Why Chipmakers Moved to ThemJul 28, 2026 · 10 min read
Article · aiRAG vs Fine-Tuning: Which One Your Use Case Actually NeedsJul 27, 2026 · 10 min read
Article · securityPrompt Injection: The Unsolved Security Hole in AI AppsJul 27, 2026 · 13 min read
Article · chipsThe Foundry Moat Moved to the PackageJul 27, 2026 · 5 min read
Article · aiWhat an AI Agent Really Is: Stripping Away the HypeJul 27, 2026 · 14 min read
Article · newsLinux Developers Are Pushing Anthropic to Ship an Official Claude Desktop AppJun 23, 2026 · 3 min read
Guide · aiThe best budget laptops for programming and AI work in 2026Jun 20, 2026 · 13 min read
Guide · aiThe best laptops for running local AI models in 2026Jun 20, 2026 · 12 min read
Guide · aiThe best GPUs for running large language models locally in 2026Jun 20, 2026 · 10 min read
Guide · aiThe best mini PCs for local AI inference in 2026Jun 20, 2026 · 10 min read
Article · newsApple Renames and Rebuilds Siri as 'Siri AI': Powered by Google on the Back EndJun 19, 2026 · 3 min read
Article · aiOpenAI Acquires Ona to Give Codex Agents a Persistent Home in Enterprise CloudsJun 19, 2026 · 3 min read
Guide · aiThe best AI coding assistants in 2026Jun 19, 2026 · 10 min read
Article · newsXiaomi's MiMo Code Claims to Out-Agent Claude Code on 200-Step Tasks: What the Numbers Actually ShowJun 19, 2026 · 4 min read
Article · newsFCC Waives Amazon Leo's Satellite Deployment Deadline, Clearing Path for a Starlink RivalJun 14, 2026 · 3 min read
Tutorial · aiHow to choose the right quantization for a local LLMMay 24, 2026 · 4 min read
Article · aiHow a transformer model actually worksMay 13, 2026 · 4 min read
Article · aiThe real difference between training and inferenceMay 12, 2026 · 4 min read
Article · aiWhat a context window actually isMay 11, 2026 · 4 min read
Article · aiWhat RAG actually is and is notMay 10, 2026 · 4 min read