Verbose Prompts Help Vision-Language Models Resist Image Corruption
According to a study published on ArXiv CS.AI, padding prompts makes vision-language models more robust to image corruption, whereas fine-grained questions increase model fragility.
What matters now
OpenAI News announced that OpenAI shared a framework for tracking, investigating, and disclosing model misalignment.
TechCrunch AI reports that Google DeepMind has launched a new institute designed to widen discussion and surface differing views on artificial general intelligence.
In a paper published on ArXiv CS.AI, researchers evaluated alignment midtraining up to 110-billion-parameter models and found that its steering effects are fragile against competing finetuning data.
OpenAI News reported that new OpenAI Economic Research examines how workers utilize AI beyond traditional roles and adopt new recurring activities.
Good Start Labs is using strategy games to train artificial intelligence models for real-world tasks, Latent Space reported. In a recent experiment, training a 30B model inside the board game 1830 improved its scores on the Finance-Agent benchmark when configured as a multi-turn terminal agent.
Startup AIUC announced a $40 million Series A funding round to develop AIUC-1, an AI agent standard backed by insurance, according to Latent Space. Cofounder Rune Kvist stated that building trust and resolving liability will be critical as companies deploy autonomous AI systems.
According to Latent Space, the AI Evaluator Forum has published AEF-1, establishing a baseline standard for independent third-party AI evaluations as frontier labs face growing calls for external oversight.
Loading selected videos…