rollup.
687 writings found
Page 3
Holo4 Shows Open Models Can Handle Real Business Workflows
H company's new agentic models combine GUI, API, and code interfaces in one system, challenging the notion that only frontier models can automate complex business tasks.
Google's AI Video Co-Director Solves the Long-Form Consistency Problem
Google researchers introduce a multi-agent framework that maintains visual coherence across minutes-long AI-generated videos, tackling character drift and cascading failures in generative pipelines.
Tiered Human-in-the-Loop: When HITL Speeds You Up
Why blindly adding human review to every AI decision slows systems down. How to architect HITL for speed, safety, and continuous learning.
Holo4: The Multi-Interface Agent That Actually Works
New agentic models that seamlessly switch between GUIs, APIs, and code. Why this matters for real business automation.
How SIG Apps is Reshaping Kubernetes for AI and Beyond
Inside the Special Interest Group maintaining Kubernetes core workload APIs, and why their evolution matters for the future of cloud infrastructure.
DSpark Vision Drafters Bring 3x Speedups to Edge VLMs
Speculative decoding for vision-language models achieves significant inference speedups on edge devices and GPUs with minimal parameter overhead and day-one framework support.
Speculative Decoding Brings VLMs to the Edge
Liquid AI's DSpark draft model achieves 3x faster vision-language inference on-device and 2.66x on H100s while adding just 8.9% parameters.
Claude Opus 5.5 and GPT-6 Luna: The Price War Gets Serious
OpenAI and Anthropic released major model updates with aggressive pricing cuts. What does this mean for your applications and the future of AI infrastructure?
Professional Skepticism in an AI-Driven Development World
Exploring how developers can apply critical thinking, test-driven development, and rigorous testing practices to build reliable AI systems and agents.
Speculative Decoding Brings VLMs to the Edge
Liquid AI's DSpark drafter achieves 3.13x speedups on-device by applying speculative decoding to vision-language models, with day-one support for llama.cpp and MLX.
Skepticism as a Superpower in the Age of AI Testing
Professional skepticism matters more than ever. David Burns on test-driven development for AI agents, flaky tests, and building reliable systems.
Meta's Muse Exposes a New Model for AI Computers
Meta's Muse VM deliberately exposes its filesystem, marking a philosophical shift in how AI platforms should work compared to ChatGPT and Gemini.
Claude Opus 5.5 and GPT-6 Luna Start a Price War
Anthropic and OpenAI released new models today with aggressive pricing. Here's what it means for building AI applications.
Open Source Methodologies Are Reshaping How Enterprises Build Software
Open source practices are moving beyond hobby projects into enterprise development. Here's why that matters for how teams ship code at scale.
Speculative Decoding Meets Vision: LFM2.5-VL Gets the Speed Treatment
Liquid AI's DSpark draft model accelerates vision-language inference by 2.3x on-device and 20x on GPU, but Amdahl's law reveals the real bottleneck.