DatologyAI builds tools to automatically select and optimize the best data on which to train AI models, leading to better, smaller models which train faster.
Today, we’re excited to launch the Datology Curation Studio, customizable frontier data curation for every team training its own models.
Data quality is the ultimate compute multiplier, and Curation Studio makes it easy for everyone.
New from DatologyAI: Zephon, an open-source data loader for text and multimodal training.
Resume a run on a different number of GPUs and most loaders quietly change the data. In our tests, GPU count alone moved eval scores by up to 0.82 points, about 4x the gap used to choose
Dumplings & Data at SF Tech Week today, with @momolicious.
Thanks to everyone who lined up for momos and stayed to talk data. Models are what they eat.
Big week continues....
Tomorrow morning: something new for anyone loading data onto GPUs.
Friday: the big one.
@arimorcos closed his keynote at @AIconference last week with a customer story: Thomson Reuters.
Decades of proprietary data, deep legal expertise, and frontier token bills eating into margins. What they did, in four clips.
Better data. Better models. A big week ahead.
We’re excited to share what we’ve been building at DatologyAI, with some exciting product updates and a few opportunities to meet the team behind the research.
Tuesday, 10/6: Join us for Dumplings & Data at SF Tech Week. We’ll also