There are PetaBytes of important scientific data locked in archival file formats. The first step is to make this efficiently readable.
I’ve been reading this sentiment on HN since GPT4o, yet models got better and better
LLMs will obviously keep getting incrementally better, but in order to get the kind on jump that LLMs themselves were, the sentiment is that we need something more.
[1] https://www.synbiobeta.com/read/anthropic-is-hiring-biologis...
Sounds like the oil scare from 90' - we thought we were gonna run out of oil. But as oil gets more expensive it pays to dig further down to find the stuff that didn't make sense to dig up before.