Using Gemma 4, Ollama, OpenAI Agents SDK, and Tavily MCP to build a lightweight research agent The post From Local LLM to Tool-Using Agent appeared first on Towards Data Science .
By Shuai Guo
Understanding ow LLMs interact with the world around them, from returning data to taking action The post Tool Calling, Explained: How AI Agents Decide What to Do Next appeared first on Towards Data Science .
By Maria Mouschoutzi
A hands-on walkthrough of code execution with the OpenAI Agents SDK and Docker The post Build an LLM Agent That Can Write and Run Code appeared first on Towards Data Science .
By Shuai Guo
The article discusses a specialized LLM inference runtime designed for real-time applications, such as a 33 ms robot control cycle. Unlike typical runtimes that ignore physical deadlines, this system refuses new requests when the deadline is at risk, evicts key‑value cache entries based on meaning rather than age, and is implemented entirely in hand‑written CUDA without relying on cuBLAS or libtorch.
By Anubhab Banerjee
To help robots do chores in places like homes and factories, a new approach from MIT uses one language model to clarify users’ instructions, then another to ignore irrelevant info.
By Alex Shipps | MIT CSAIL