the log · 132 entries
#infrastructure
Every entry filed under Infrastructure.
all 132
ai-machine-learning 13
ai-engineering 71
infrastructure 97
machine-learning 3
programming 51
software-development 28
technology 96
technology-leadership 29
tutorials 1
2026.08.247 min read
MCP's Stateless Release Candidate Removes a Deployment Burden
MCP's July 28 release candidate removes the session handshake and puts metadata on each request. That simplifies deployment, but applications still need to manage their own state.
2026.08.238 min read
Cloudflare's Kitesurf Offers a Lighter Browser for AI Agents, With Tradeoffs
Cloudflare says Kitesurf uses up to 7 times less memory than Chromium for agent browsing tasks. Its Rust and WebAssembly design saves resources, but latency, JavaScript support, and session…
2026.08.226 min read
GitLab Exploitation Shows Why Critical Patches Can't Wait for Scheduled Maintenance
Exploitation of a GitLab flaw was observed within hours of disclosure. The contrast with a centrally patched Entra ID flaw shows the response capacity that self-managed infrastructure…
2026.08.217 min read
Google’s Marvell Deal Ties Its Cloud Plans More Closely to Chip Design
Google’s right to buy up to $12.2 billion in Marvell shares links chip purchases to equity. The arrangement could give Google more control over its AI infrastructure costs and less…
2026.08.197 min read
Etched's Jane Street Deployment Puts Its Inference Chips to a Production Test
Etched says its first rack is running production workloads at Jane Street. The deployment gives its inference hardware an early test, but software integration and manufacturing scale remain…
2026.08.177 min read
Data Center Operators Pursue IPOs as AI Demand Pushes Valuations Higher
Vantage, Switch, CyrusOne and DayOne signaled public listing plans in one week. AI demand is lifting their expected valuations, but power constraints, operating costs and investor caution…
2026.08.167 min read
AI API Prices Are Falling Fast. Model Routing Is Becoming Core Infrastructure.
Rapid AI price cuts make a fixed model choice expensive to maintain. Task-aware routing offers a way to reduce costs while keeping harder work on more capable models.
2026.08.156 min read
Qwen 3.8-27B Gives Self-Hosted Agents a Stronger Case
Alibaba's Qwen 3.8-27B combines Apache 2.0 licensing, single-GPU deployment and stronger reported agent benchmarks. The next test is how those gains hold up in local workloads.
2026.08.147 min read
Cerebras Reports 750 Tokens Per Second for GPT-5.6 Sol Ultrafast
Cerebras and OpenAI's Ultrafast tier promises up to 750 output tokens per second. On-chip memory helps explain the speed, though production performance and capacity remain open questions.
Nothing on this page matches. Press Enter to search every entry.