TechCrunch AI reports that Anthropic is now operating a lab where it conducts biology experiments. The company has positioned AI as a possible tool for curing human disease but also warns of potential risks.
According to TechCrunch AI, UP.Labs—now operating as Vantora—has secured $100 million to develop new startups aimed at industrial corporations, with a focus on physical AI technologies.
TechCrunch AI reports that Vals, a company backed by Andreessen Horowitz, has launched with the goal to become a reliable and neutral platform for AI benchmarking. Vals positions itself as a trustworthy resource as the number of AI models continues to grow.
TechCrunch AI reports that Google DeepMind has launched a new institute designed to widen discussion and surface differing views on artificial general intelligence.
According to a study published on ArXiv CS.AI, padding prompts makes vision-language models more robust to image corruption, whereas fine-grained questions increase model fragility.
In a paper published on ArXiv CS.AI, researchers evaluated alignment midtraining up to 110-billion-parameter models and found that its steering effects are fragile against competing finetuning data.
Good Start Labs is using strategy games to train artificial intelligence models for real-world tasks, Latent Space reported. In a recent experiment, training a 30B model inside the board game 1830 improved its scores on the Finance-Agent benchmark when configured as a multi-turn terminal agent.
Startup AIUC announced a $40 million Series A funding round to develop AIUC-1, an AI agent standard backed by insurance, according to Latent Space. Cofounder Rune Kvist stated that building trust and resolving liability will be critical as companies deploy autonomous AI systems.
According to Latent Space, the AI Evaluator Forum has published AEF-1, establishing a baseline standard for independent third-party AI evaluations as frontier labs face growing calls for external oversight.