Towards Data Science

The Problem with pandas Isn’t Performance. It’s Cognitive Overhead.

Faster dataframe engines are nice, but they don't reduce the amount of syntax an analyst has to hold in their head. The post The Problem with pandas Isn’t Performance.

Towards Data Science
4d ago

AI Made Data Scientists Faster. Now It’s Expanding the Job.

AI has accelerated data scientists’ productivity, but its influence extends beyond speed. The technology is reshaping who owns data, how judgment is exercised, and the overall career trajectory of data scientists. These changes signal a broader transformation in the field’s structure and responsibilities.

By Yu Dong
Simon Willison
Sep 20

Quoting voxium

Simon Willison describes his experience at a large company where all documentation, code, tests, PRDs, tickets, and reports are generated by Claude Code. His team is forced to ship rapidly, working long hours, yet management insists that code push is not a bottleneck, leading to frustration and a lack of meaningful reading or review. The situation highlights a reliance on AI-generated content that may undermine quality and collaboration.

Towards Data Science
Aug 28

Why Claude Code Time Estimates Are Poor

The article titled "Why Claude Code Time Estimates Are Poor" discusses the challenges and shortcomings of using Claude, an LLM, for estimating code development time. It highlights how these estimates can be unreliable and offers insights into improving communication when working with LLM programming tools.

By Eivind Kjosbakken
Towards Data Science
Aug 20

Three Kinds of RAG Corpus, and What It Costs to Build for the Wrong One

The article explains that enterprise document intelligence can be categorized into three distinct corpus types, each requiring a specific architecture. It outlines how to determine the shape of a document collection through three key questions. The piece also discusses the costs associated with building a system for the incorrect corpus type.

By angela shi