Researchers have unveiled *-PLUIE, a novel evaluation metric for AI text. By measuring an LLM's confidence on 'Yes/No' answers, it offers a faster, more computationally efficient alternative to traditional LLM-as-a-judge methods.
Research WireDoes AI fine-tuning impart new skills or just unlock existing knowledge? A new research paper proposes 'task complexity' as a formal metric to finally settle the debate.
Research WireA breakthrough AI method, DMTS-NC, dramatically accelerates molecular simulations. By teaching a simple model to mimic a complex one, this new approach could revolutionize drug discovery and materials science by making simulations faster and more efficient.
Research WireEver wonder how a new app knows what you like? Researchers have developed a 'training-free' method to solve this 'cold-start' problem, using world knowledge to ask smarter questions and personalize experiences faster than ever.
Research WireAstronomers face a major challenge: data from different telescopes isn't always compatible. A new study reveals how simple AI models, pre-trained on vast low-resolution datasets, can be effectively adapted to analyze higher-quality data from newer surveys.
Research WireResearchers have unveiled MacroGuide, an AI that intelligently designs complex, ring-shaped macrocycle molecules. This breakthrough could accelerate drug discovery for tough diseases.
Research WireResearchers have successfully fine-tuned a physics-aware AI foundation model, Poseidon, to create a highly accurate weather emulator for Mars. This breakthrough could revolutionize planetary science and mission planning.
Research WireResearchers from Google have unveiled BPP, a new method that gives robots a long-term memory. It helps them focus on key past events to solve complex, multi-step tasks.
Research WireA new paper challenges the complex-by-design approach for AI in science. By first standardizing data like molecules into a 'canonical' form, researchers can use simpler diffusion models, bypassing the need for specialized equivariant architectures.
Research WireA new benchmark, Mage-Bench, tests the strategic reasoning of top AI models by forcing them to play Magic: The Gathering, a game of immense complexity, hidden information, and long-term planning.
Research WireResearchers have unveiled a new AI model that infers the underlying reasons for your online activity to deliver hyper-personalized news recommendations across domains.
Research WireResearchers have unveiled the first-ever scaling laws for discrete diffusion language models, discovering a novel training method that makes them 12% more efficient.
Research Wire