OpenAI unveiled a 750-task assessment to evaluate whether AI systems can support realistic life science research, with the top model achieving only a 36.1% pass rate. Performance declined notably when handling documents and complex datasets versus text-only tasks, showing AI cannot yet autonomously conduct scientific research.