RSS.Social

Blogs on Data Engineering Blog & Second Brain

follow: @[email protected]

Posts

If You Always Enjoy It, You're Not Pushing It Hard Enough—Benn Stancil (Show Your Workflow #1)

Writing really is an Emotional Rollercoaster

Figma for Agents: How Airflow's Creator Coordinates AI ft. Maxime Beauchemin

The Act and the Outcome of Creation

The Grammar of Data: Define Once, Run Anywhere with Cross-Engine Expressions

Where AI Agents Belong in Data Engineering: The Correctness Layer

The Process of Smart Note-Taking

Operationalizing Data Orchestration: Best Practices for DevOps, Infra, and Code Locations

Vibe Coding Is Dangerous, Agentic Engineering Isn't—Wes McKinney

Beyond the Semantic Layer: Building a Context Layer for the Agentic Era

Plan Mode All the Time, Substrait over SQL, and the End of the DE Role ft. Chris Riccomini

The Dagster Almanack: From Complexity to Composability

Internal vs. External Storage? What's the Limit of External Tables

AI Reveals Why BI Still Matters

Specs Over Vibes: Consistent AI Results ft. Mark Freeman

Building an Agent-Friendly, Local-First Analytics Stack with MotherDuck and Rill

Why I Still Blog — and Why the Future of Blogging Is Connected

Git for Data Applied: Comparing Git-like Tools That Separate Metadata from Data

Building an Obsidian RAG with DuckDB and MotherDuck

Arch Linux (Omarchy) — 8 Months Later: The Good, the Bad, and the Fixable

Why Coinbase and Pinterest Chose StarRocks: Lakehouse-Native Design and Fast Joins at Terabyte Scale

A Diary of a Data Engineer

Well Being in Times of Algorithms

Opinionated Data Platforms vs. Open-Source: The Chef’s Choice for Your Data Platform

Simplicity of a Database, but the Speed of a Cache: OLAP Caches for DuckDB

Dlt+ClickHouse+Rill: Multi-Cloud Cost Analytics, Cloud-Ready

Branch, Test, Deploy: A Git-Inspired Approach for Data

Multi-Cloud Cost Analytics: From Cost-Export to Parquet to Rill

Boredom is the New Luxury

4 Senior Data Engineers Answer 10 Top Reddit Questions

Data Modeling for the Agentic Era: Semantics, Speed, and Stewardship

Beyond Basic ETL: Enterprise Data Capabilities Without the Complexity

Why I Don't Research; and Write from Experience

Data Modeling Guide for Real-Time Analytics with ClickHouse

Why Semantic Layers Matter—and How to Build One with DuckDB

My Journey from macOS to Arch Linux with Omarchy

Summer Data Engineering Roadmap

Why Are We Here on Earth? True Happiness, Giving Up Control, or the Trap of Worshipping Earthly Things?

The Data Engineering Toolkit: Infrastructure, DevOps, and Beyond

Has Self-Serve BI Finally Arrived Thanks to AI?

Universal Data Orchestrator in Action: Enterprise Best Practices

The Open Lakehouse Stack: DuckDB and the Rise of Table Formats

Self-Host & Tech Independence: The Joy of Building Your Own

The Open Table Format Revolution: Why Hyperscalers Are Betting on Managed Iceberg

Configure, Don't Code: How Declarative Data Stacks Enable Enterprise Scale

The Data Engineer's Guide to Efficient Log Parsing with DuckDB/MotherDuck

The Universal Data Orchestrator: The Heartbeat of Data Engineering

What «Shifting Left» Means and Why it Matters for Data Stacks

Vector Technologies for AI: Extending Your Existing Data Stack

Scaling Beyond Postgres: How to Choose a Real-Time Analytical Database

A Beginner’s Guide to Geospatial with DuckDB

Finding Flow: Escaping Digital Distractions Through Deep Work and Slow Living

Why Pivot Tables Never Die

The Data Engineering Toolkit: Essential Tools for Your Machine

Designing a Declarative Data Stack: From Theory to Practice

Semantic Layer and AI: The Future of Data Querying with Natural Language

Universal Semantic Layer: Capabilities, Integrations, and Enterprise Benefits

Building with Bluesky: Inside the New Open Social Network

Exploring the Semantic Layer Through the Lens of MVC

15+ Companies Using DuckDB in Production: A Comprehensive Guide

BI-as-Code and the New Era of GenBI

The Enterprise Case for DuckDB: 5 Key Categories and Why Use It

The Rise of the Declarative Data Stack

My Obsidian Note-Taking Workflow

The Plain Text Workflow: How Vim and Markdown Became My Backbone

Data Modeling - The Unsung Hero of Data Engineering: Architecture Pattern, Tools and the Future (Part 3)

Data Modeling – The Unsung Hero of Data Engineering: Modeling Approaches and Techniques (Part 2)

Data Modeling – The Unsung Hero of Data Engineering: An Introduction to Data Modeling (Part 1)

Pandas 2.0 and its Ecosystem (Arrow, Polars, DuckDB)

Finding My Pathless Path

Modern Data Stack: The Struggle of Enterprise Adoption

Why Vim Is More than Just an Editor – Vim Language, Motions, and Modes Explained

The Open Data Stack Distilled into Four Core Tools

Data Integration as Code: Configuring Airbyte and dbt with Python (Dagster)

Rust for Data Engineering

The Rise of the Semantic Layer

Data Lake / Lakehouse Guide: Powered by Data Lake Table Formats (Delta Lake, Iceberg, Hudi)

Data Orchestration Trends: The Shift From Data Pipelines to Data Products

Personal Knowledge Management Workflow for a Deeper Life — as a Computer Scientist

Building an Analytics API with GraphQL: The Next Level of Data Engineering?

How to Take Notes in 2021?

Saying Goodbye to WordPress, a Homecoming

Building a Data Engineering Project in 20 Minutes

Business Intelligence meets Data Engineering with Emerging Technologies

Email, and the way we (should) communicate at work

Open-Source Data Warehousing – Druid, Apache Airflow & Superset

OLAP, what’s coming next?

Today’s Office – The Location Independent Lifestyle

Tools I Use – Part III

Data Engineering, the future of Data Warehousing?

Data Warehouse vs Data Lake | ETL vs ELT

Tools I Use – Microsoft OneNote – Part II

What Data Warehouse Automation tools are on the market

Why automate? What does Data Warehouse Automation for us?

Why Data Warehouse Automation is not more popular

Data Warehouse Automation (DWA) – Series

Tools I Use

The more you share the more you get..

Migrate from Oracle to Microsoft (Views) – Part III

Migrate from Oracle to Microsoft (Views) – Part II

Migrate from Oracle to Microsoft (Views) – Part I

SSAS Cubes – Dynamic generation of partition

Need any Headphones?

Do you listen to music while working?

Annoying advertising, what you can do

Annoying advertising, what you can do

ORDER BY Oracle vs. Microsoft SQL