Building a Data Lakehouse with DuckDB and DuckLake
Read the original on Towards Data Science →The Flow has not summarised this story yet — read it at Towards Data Science.
The Flow has not summarised this story yet — read it at Towards Data Science.
A small experiment in remote SQL execution The post Running SQL Concurrently Across Three Remote DuckDB Servers with Quack appeared first on Towards Data Science .
The article recounts the author's experience of moving a Dockerized data pipeline from a local laptop to AWS, highlighting the challenges that arose when the environment changed. It explores lessons learned about container behavior, networking intricacies, and hidden assumptions that were previously taken for granted in a local setup. The post serves as a practical guide for developers facing similar transitions to cloud infrastructure.
This is the opening piece of a four-part deep dive series, on building a high-frequency streaming pipeline against a live public API. The data source is openSenseMap, a citizen-science IoT network used for climate research, mostly in Germany.
A hands-on walkthrough of a hybrid local-cloud workflow using Gemma 4 and GPT-5. 4, with reasoning and structured outputs The post Stop Choosing Between Local and Cloud LLMs: A Field Guide to Hybrid Patterns appeared first on Towards Data Science .
The article explains how to connect a LangGraph AI agent to a Postgres database. It covers running the backend locally using Docker and deploying it in the cloud. The guide provides practical steps for setting up the database connection and managing the agent’s data storage.