abundance and unrest collide.webp

Abundance and Unrest Hit the Middle of the Labor Market · Hanh D. Brown

Short answer Which workers does AI hit first, and what does the forecast actually say? White-collar work goes first. Blue-collar work goes later. The industrial-era story has the order wrong. Anything that lives in bits gets automated now. Anything that needs hands on atoms takes longer. Abundance and unrest arrive together in three to seven …

Abundance and Unrest Hit the Middle of the Labor Market · Hanh D. Brown Read More »

1787858702320 n0xwek

Disaggregation Is a Thousand-GPU Problem

Every major inference framework shipped prefill-decode disaggregation this year. NVIDIA built it into Dynamo. SGLang made it the default for large-scale deployments. vLLM added a KV connector API to support it natively. The consensus is forming fast: split your prefill and decode onto separate GPU pools, and throughput improves. The consensus is wrong for most …

Disaggregation Is a Thousand-GPU Problem Read More »

ai economic future weak links.webp

Slow Upside, Fast Downside, Weak Links · Hanh D. Brown

AI’s Economic Future: Slow Upside, Fast Downside, Weak Links. A chain breaks at one link and strengthens one at a time. That asymmetry is why AI’s economic upside arrives slowly and its downside arrives fast. ai-economics ai-policy ai-careers A chain is only as strong as its weakest link. Artificial Intelligence (AI) makes the strong links …

Slow Upside, Fast Downside, Weak Links · Hanh D. Brown Read More »

Rosidi AI Data Analysis Mistakes 1

I Asked ChatGPT to Analyze 3 Datasets. It Made the Same Mistakes Every Time

We ran an experiment: three small datasets, one AI model, and the questions a business team asks in a normal week — what’s our average delivery time, which region is our best performer, how many athletes are in this file. Then we added a review pass. We handed the model its own answer back and …

I Asked ChatGPT to Analyze 3 Datasets. It Made the Same Mistakes Every Time Read More »

ai reinvention of computing.webp

AI Is Not a Faster Chip. It Is a Reinvention of Computing. · Hanh D. Brown

AI Is Not a Faster Chip. It Is a Reinvention of Computing.. AI is not a faster chip. It is a wholesale reinvention of how computers work, and the sentence the chip-and-forecast discussion does not say is the one that matters most. ai-hardware ai-economics ai-policy For sixty years, software was written by humans typing rules …

AI Is Not a Faster Chip. It Is a Reinvention of Computing. · Hanh D. Brown Read More »

agentic video keyword blog header.width 1300

Introducing Agentic Video in Gemini

Today, we’re launching agentic video understanding across our latest models: Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite. This new capability improves accuracy while dramatically reducing token usage and costs for video analysis. Similar to agentic vision, which combines code execution with Gemini models’ native image understanding, agentic video understanding uses Gemini’s native video tools …

Introducing Agentic Video in Gemini Read More »

ai blind spot.webp

The AI blind spot that nobody in the room will name · Hanh D. Brown

The AI blind spot that nobody in the room will name. AI vendor pitches start with the model. The model is real and not the differentiator. Three numbers tell the rest of the story: 367, 6 versus 1, and 95. ai-strategy ai-and-work ai-markets Every Artificial Intelligence (AI) vendor pitch a chief information officer has heard …

The AI blind spot that nobody in the room will name · Hanh D. Brown Read More »

ai optimism grammar.webp

Four Tells in Every Announcement · Hanh D. Brown

AI Hype vs Reality: Four Tells in Every Announcement. AI optimism in 2026 has a grammar. Forecasts read as facts. Product lists pass for forecasts. Four tells let you read any announcement honestly. reading-ai-news ai-markets ai-policy Artificial Intelligence (AI) optimism in 2026 has a grammar. Forecasts arrive dressed as descriptions. Product lists arrive dressed as …

Four Tells in Every Announcement · Hanh D. Brown Read More »

kdn speed up llm inference with dspark speculative decoding feature

Speed Up LLM Inference with DSpark Speculative Decoding

There are many ways to get more from the models and GPU infrastructure you already have. Quantization, optimized kernels, and better inference engines can all help, but speculative decoding is especially useful because it can increase generation speed without simply adding more GPUs. There are now several approaches to speculative decoding. Traditional methods use a …

Speed Up LLM Inference with DSpark Speculative Decoding Read More »

Scroll to Top