Patent x AI Tech Media
Technical insights on GPU inference, local LLMs, and developer tools
Latest Tech News
Anthropic Fable 5 Updates; GitHub Copilot API Adds Agent Activity for Claude/Codex
Anthropic announces significant updates to its Fable 5 model or system, including a redeployment and enhanced biology sa...
SQLite & DatabasesPostGIS 3.7.0beta2, SQLite VM Optimizations Lead Database News
PostGIS launches its 3.7.0beta2, bringing new features and PostgreSQL 19 compatibility to geospatial users. Concurrently...
Rust, Cloudflare & Dev StackDocker Sandboxes for AI, Cloudflare Radar Researcher & GitHub Enterprise Apps
This week's top releases feature Docker Sandboxes for securely running AI agents and Cloudflare's new AI-powered Radar R...
GPU, CUDA & Autonomous DrivingPyTorch Metal Codegen Adds Uint Support; NVIDIA Blackwell/Rubin AI Factory Launched; AMD ROCm LLM Guide
This week features an official PyTorch release fixing `uint16` support for Metal codegen, a major NVIDIA AI factory laun...
Local AI & Open Modelsllama.cpp, PyTorch Updates Boost Local Inference; New MoE Model Trends
Today's top stories feature significant updates to core local AI libraries: `llama.cpp` enhances WebGPU acceleration and...
Cloud AI, APIs & MCPClaude Code v2.1.226 Released, Google Details AI Agent Scaling & Verifiable Research Frameworks
Anthropic has updated its Claude Code library to v2.1.226, focusing on bug fixes and reliability for developers. Meanwhi...
Latest Deep Dives
My RAG's Model Had Already Read the Books - Per-Book Verdicts, a Fake Regression, and Catching a 9B Leaking Prior Knowledge
I run a local RAG over all seven Harry Potter novels with a 9B model that knows the franchise cold. Re-running my eval o...
GPU & InferenceSolving 7×7 Killall-Go Opening JA on a Single RTX 5090
We reproduce the NeurIPS 2023 online fine-tuning Killall-Go solver natively on a single RTX 5090 and prove the 7x7 openi...
GPU & InferenceSolving Cho Chikun Life-and-Death Problems on a Single RTX 5090
A 2x2 cross-generation benchmark (hardware x algorithm) of the Relevance-Zone life-and-death solver on Cho Chikun's prob...
GPU & InferenceThe Same RTX 5090, but the GPU Sat Idle — a CPU-Bound Go Solver and the Case for L2 Cache
Second in the RTX 5090 series — but the GPU sits ~23% idle. This Go life-and-death solver is CPU-bound, and the prime su...
GPU & InferenceOne RTX 5090 vs a 12-GPU Cluster — Benchmarking a Decade of GPUs on the Same Go Proof
A 2023 paper solved a tiny 7x7 Go position on a 12-GPU cluster. I re-ran the exact same solver on a single RTX 5090 — id...
GPU & InferenceINT8 Q/DQ Calibration on Blackwell: 1.8× the TRT 10 + FP16 Baseline
A practical walkthrough of doing INT8 post-training quantization the right way on RTX 5090 + TensorRT 11. 1,500 stratifi...