Skip to content
GenX AI Tools

Expert reviews and comparisons of the best AI tools for work, creativity, and business.

  • Home
  • Articles
  • About
  • Contact
  • Privacy Policy
  • Terms
  • facebook.com
  • twitter.com
  • t.me
  • instagram.com
  • youtube.com
Subscribe

Posts by Tai Nguyen

Home » Archives for Tai Nguyen
About Tai Nguyen
Tai Nguyen writes GenX AI Tools, a single-operator publication on AI models, inference engines, and the hardware that runs them. Every claim is checked against a primary source before publication.
Consumer GPU on a dark workbench fed by an amber power cable through a plug-in power meter, illustrating tokens per watt under a power cap
Posted inLocal LLMs & Hardware

Tokens per Watt: What Power Capping Actually Costs a Local LLM Build

Throughput per GPU was the number that mattered in AI infrastructure for most of 2024 and 2025. It has quietly stopped being the number that matters. On September 16, Emerald…
Posted by Tai Nguyen October 11, 2026
Two compact desktop AI machines linked by a 200G QSFP56 cable as a two-node home inference cluster
Posted inLocal LLMs & Hardware

NVIDIA DGX Spark 64GB: Half the Memory, Higher Price, and the Home Cluster That Actually Matters

NVIDIA will sell a 64GB DGX Spark starting October 23, 2026, through Acer, ASUS, Dell, Gigabyte, HP and MSI, from $4,999. It carries half the unified memory of the 128GB…
Posted by Tai Nguyen October 10, 2026
Silicon wafer and stacked HBM memory dies beside DDR5 DIMM modules, illustrating DRAM capacity diverted to AI memory
Posted inLocal LLMs & Hardware

HBM Is Eating Your RAM: What a 3x DRAM Price Means for Local LLM Builds

The reason the RAM in your local LLM box costs three times what it did a year ago is not that anyone started buying more RAM. It is that three…
Posted by Tai Nguyen October 9, 2026
800V DC busway and liquid cooling serving high-density AI racks
Posted inAI Infrastructure & Data Centers

800V DC and the Cooling Wall: How 120 kW Racks Rewrote Data Center Power

Data center power engineering has spent thirty years optimizing a chain nobody outside the electrical room sees: medium-voltage AC at the fence line, transformers, switchgear, UPS, power shelves, then a…
Posted by Tai Nguyen October 8, 2026
Strata inference engine: a 125B mixture-of-experts model split across GPU, RAM, CPU and SSD on a gaming PC
Posted inLocal LLMs & Hardware

Strata: The Inference Engine That Runs a 125B MoE on a Ryzen 5900X and an RTX 3090

Strata is not a model. It is an inference engine, and the distinction matters: the 125-billion-parameter Qwen3.8-Flash-Next it runs would not fit on a gaming GPU by a wide margin,…
Posted by Tai Nguyen October 8, 2026
AMD Helios rack-scale accelerator: 31TB of HBM4 stacked in a single data center cabinet
Posted inAI Infrastructure & Data Centers

AMD Helios Ships: The $1.2B Vultr Order That Puts 31TB of HBM4 in One Rack

AMD Helios Ships: Vultr's $1.2B First Rack Order Puts 31TB of HBM4 on the Grid On September 30, HPE announced a $1.2 billion order from Vultr to deploy AMD Helios…
Posted by Tai Nguyen October 7, 2026
NVIDIA RTX Spark superchip powering a compact Windows AI desktop with 128GB unified memory
Posted inLocal LLMs & Hardware

NVIDIA RTX Spark: The 1-Petaflop Windows AI Superchip Ships This Month

NVIDIA's RTX Spark PCs start reaching store shelves this month, and Microsoft is putting the spotlight on them at a joint Windows and Surface event in San Francisco on October…
Posted by Tai Nguyen October 7, 2026
Terminal coding agent harness routing requests between planning and execution model nodes
Posted inAI Agents & Automation

Pi 1.0: The Minimal Coding-Agent Harness That Routes Planning to Opus and Execution to GPT

Pi 1.0 shipped on October 1, 2026, and it is the most interesting coding-agent release of the year precisely because it added features by refusing to add most of them.…
Posted by Tai Nguyen October 6, 2026
Sparse MoE neural network with glowing nodes representing Reflection Beam 501B architecture
Posted inAI Models & Releases

Reflection Beam 501B: Open-Weight MoE Announced with Apache 2.0 Release Planned

Reflection AI unveiled Beam on October 5, 2026 — a 501 billion parameter sparse Mixture-of-Experts model with 23 billion active parameters per token, aimed at coding, reasoning and agentic workloads.…
Posted by Tai Nguyen October 6, 2026
Neural network decision node glowing blue with branching paths
Posted inAI Models & Releases

AWS Strands Decider 2B: Open 2B Decision Model Released for Local Agents

Amazon Web Services quietly shipped a new class of model on October 1, 2026. Strands Decider 2B is not a chat model. It is a 2-billion-parameter decision model built on…
Posted by Tai Nguyen October 5, 2026

Posts pagination

1 2 3 Next page
Footer
GenX AI Tools is a single-operator publication on AI models, inference engines, and the hardware that runs them. Some product links are affiliate links: if you buy through them the site may earn a commission at no cost to you. About | Contact | Privacy Policy | Terms.
Copyright 2026 — GenX AI Tools. All rights reserved. Bloghash WordPress Theme
Scroll to Top