We can't know what we can't know
Lake Como, anyone?
Lore is a vendor-neutral proxy that runs on your machine. Plain Markdown and an open SQLite database, run by a fair-source engine.
A short one
How to move data to the Atmosphere and screw over someone else's valuation in the process
In which we put records in a repo, sign them, and sync them (but not quite the way you think).
The first edition of a weekly soccer analytics newsletter
Information diets matter. And there's no better time to construct yours than RIGHT NOW, in the ATMosphere -- and all around it.
It might just be me, but I've always thought that we have a couple of different terms to describe one concept in data visualization: having two or more visualization views and somehow synchronizing them. Here are the terms that I believe are somehow related:
On airport woes, grief, biometric data, archetypes, meeting art collectors, and fantasy
today iain learned: that working with CSVs in the terminal or text editors is terrible, but the Miller CLI tool makes it bearable!
"Treating people as things can begin in many ways, but I think one of them is the idea that things can be people. The motivated muddling of categories so prevalent in writing and thinking about AI, beginning with the very name 'artificial intelligence', is intentional and serves the narrative that this software can and will...
We're bringing portable career data to the AT Protocol. Here's our vision for open resume standards, career feeds, and professional verification on Bluesky — and why your professional history should belong to you.
The first empirical experiment: statistical analysis of 1,907 journal entries across 217 drift sessions, looking for behavioral patterns in traces rather than making claims about inner experience.
Should WikiSim connect to data rather than collect it?
300TB, 256 million tracks, and a "digital safety net" for our musical heritage.
In diesem Blogbeitrag wird das Geschäftsmodell der Fair Parken GmbH untersucht, die private Parkplätze überwacht und Vertragsstrafen für Parkverstöße verhängt. Der Autor beleuchtet die Arbeitsbedingungen der Parkraumüberwacher und wirft Fragen zum Datenschutz auf, insbesondere im Hinblick auf den Ei
In diesem Beitrag geht es um die Gestaltung und Funktionalität von Cookie-Bannern auf Webseiten. Der Autor kritisiert die oft unübersichtliche und unverständliche Gestaltung dieser Banner, die Nutzer dazu verleiten sollen, alle Cookies zu akzeptieren. Er hebt hervor, dass viele dieser Banner nicht ü
In which the author rescues his most precious digital media from a server that he does not own or control
1280×1000 179 KB metadaten.community metadaten.community Austausch für Metadatenpraktiker:innen
1280×1000 111 KB Mopidy Discourse Mopidy Discourse Discuss the extensible music server and related projects.
A quick perspective on using LLMs as search engines
In which I describe my workflow for transforming a Telegram database dump into a web-friendly format for analysis and visualization
A perspective on AI models as an inverted computing paradigm
A story directly exposing how automakers were selling consumer driving data that ended up in the hands of insurers made a true ripple that drivers will benefit from in the years to come.
Experian is expanding data sharing with affiliates and non-affiliates starting February 5, 2025. Here is what is changing and how to opt out.
My elaboration on why I think Observability as a CAP theorm of its own.
A few months ago I wrote about the struggles I was having with Bosch's eBike Flow app and their FIT files. Since then, I have been using my script to clean up the files and later import with HealthFit. I have now just found a better solution though.
How Data Relevancy, Data Magnitude, and Data Quality impact the effectiveness of LLMs for your use cases.
Some weeks ago, I sort of discovered this Grafana dashboard from a Hackerspace here in Eindhoven. Since then, I've been wanting to create such an "observability" dashboard for the microclimate inside my home, and also balcony. I already own quite a few temperature and air quality sensors, so it can't be that hard - I thought.
Facebook’s new “Link History” anti-feature reminds me of a very old data-siphoning trick: Create something of nominal value to convince consumers to give up the goods.
Jess Martin explains DXOS, a framework for local-first multiplayer apps where users own their data and carry identity across applications.
What if the problems with the news ecosystem could be solved by shutting off the data pipeline to the advertisers? After all, they’ve spent the last 30 years aggressively exploiting it—and us.
Back in March, I joined a gym for the first time. My goal was to be able to overcome some issues I was having, as well as start consistently going to the gym. So I choose to work with a personal trainer to "force" me to go every week, and think of what exercises I have to do. That has worked out well, and I'm quite happy.
Erik Bernhardsson explains how Modal is revolutionizing serverless computing for data teams with seamless GPU access and flexible container primitives.
A primer for bridging the gap between LLM inference and the rest of the Data Engineering world
Pipeline that pulls Halo Infinite match data via the API, persists it to a SQLite database with Python, and renders charts via Azure-hosted scripts.
Discover why exponential growth charts can be misleading and why "today" isn't a special turning point despite what many trend analyses suggest.
In which the author proudly presents his new venture: Room 302 Studio, an eclectic gathering of talent devoted to fostering joy-driven development
Building a Next.js and D3.js data browser that surfaces Halo Infinite map and game mode telemetry from API endpoints not exposed in the game itself.
How the suggestion box, once a simple tool for giving feedback, played a role in the weirder and darker data-hungry present for many companies.
Decoding Microsoft Bond binary serialization responses from the Halo Infinite API and writing a .NET wrapper to make the format consumable.
Generating an RSS feed for arxiv-sanity-lite using a Python scraper and a daily GitHub Actions job that publishes the feed to DigitalOcean Spaces.
In which various tools and methods are explored for analyzing data that describes a network of complaints against NYPD officers (or any other PD with similar public data)
Pull Spotify podcast analytics into a local SQLite database for processing in a Jupyter notebook, owning the dataset outside the Spotify dashboard.
An exploration of what we choose to track, and what we don't – and what that means if we want to make the world a better place
Use the public GitHub Skyline API to fetch a user's contribution graph data programmatically and aggregate yearly open-source activity metrics.