On Aug 28, 2026, NVIDIA released TensorRT Model Connect, an open collection of reference implementations that turns a Hugging Face or local checkpoint into native C++ inference with two commands (trtmc build, then load-and-run), skipping the ONNX export step. It emits a versioned .bundle that runs without PyTorch or Python at runtime.
NVIDIA ships TensorRT Model Connect: Hugging Face checkpoint to C++ in two commands
AI product teams can cut deployment plumbing and serve open models in native C++ for lower latency and cost; try trtmc on your current open-weight model before building a custom serving stack.
Source: NVIDIA Developer Blog
More that helps you.
Anthropic adds auto permission mode to Claude Managed Agents
On September 10, 2026 Anthropic shipped an 'auto' permission policy for Claude Managed Agents that lets the server evaluate each agent or MCP tool call and run it, deny it, or paus…
OpenHands 1.0 ships production self-hosted coding agent at 68% SWE-bench Verified
The open-source autonomous coding agent OpenHands reached its 1.0 release on September 8, 2026, adding Docker sandboxing, built-in security policies, resource limits and a plugin s…
GitHub Copilot adds parallel agent sessions and a unified Copilot experience
In its September 4, 2026 changelog, GitHub added parallel agent sessions to the Copilot app, letting multiple AI tasks run independently in separate Git worktrees, and introduced a…
Google DeepMind launches WeatherNext 3 with hourly 5km AI forecasts and API
Google DeepMind and Google Research launched WeatherNext 3 on September 3, 2026, generating 15-day probabilistic forecasts initialized hourly at up to 5km resolution across 64 ense…
Get briefs like this tuned to you.
In the app, Founder Briefs are personalized to your country, industry and stage, and you can save the ones that matter.
See plans →