In medicine, we’re already seeing startups expand into clinical decision support tools with read-access to the entire health record, which are often *decades* long per patient. Compressing not only text, but imaging, labs/trajectory, and reasoning in the least lossy way possible is critical, and representation-level tools may be the way to do it! Check out Vishnu’s work on the subject.
I've been researching how to compress coding agent context at the representation level, working directly with models' internal vectors instead of summarizing text. So far, I've hit 2x+ compression while preserving most factual recall.