Google DeepMind has announced the world's first double-blind evaluation for AI models, aiming to prevent benchmark contamination. This approach ensures that models are tested without prior exposure to evaluation questions, maintaining the integrity of the results. Collaborating with partners such as the Singapore AI Safety Institute and OpenMined, DeepMind is testing its Gemini Flash Lite model in a secure, cryptographic environment. This development is crucial for maintaining trust in AI benchmarks, ensuring they accurately represent a model's true capabilities.
Read original
© Google DeepMindGoogle DeepMind's release of Gemini Omni 1.1 Flash marks a significant step forward in generative video technology. This update introduces creative controls and capabilities that allow developers to extend scenes, specify keyframes, and upscale videos to 4K resolution. By analyzing up to 10 seconds of prior context, the model enhances visual consistency and narrative flow, making it ideal for professional use. The ability to generate lightweight previews quickly and cost-effectively further streamlines the creative process. This release empowers developers to create more polished and controllable generative video content.
© Google DeepMindGoogle DeepMind has unveiled Gemini 3.5 Transcribe, a cutting-edge speech-to-text model that excels in handling background noise, complex jargon, and disfluency cleanup. This model is integrated into various Google products, offering developers the ability to build advanced voice capabilities through the Gemini API. With impressive word error rates and support for over 85 languages, Gemini 3.5 Transcribe sets a new standard for transcription accuracy and latency. This release marks a significant improvement over previous models, making voice interactions more natural and intuitive across Google's ecosystem.
© MIT News AIMIT researchers have introduced PottsMPNN, a machine-learning framework that enhances protein design by focusing on the sequence-energy landscape rather than mimicking native sequences. This approach allows for the creation of novel proteins with structures that don't resemble any found in nature, potentially revolutionizing biological engineering. By incorporating physical principles and evolutionary information, PottsMPNN improves the prediction of protein stability and the effects of mutations. This advancement could lead to significant breakthroughs in designing proteins for diverse applications, marking a shift in how AI is used in biological research.
© Google Research BlogGoogle Research has unveiled the Planetary Prediction Engine (PPE), a groundbreaking tool within Google Earth AI that automates the entire geospatial modeling workflow. This innovation addresses the fragmented data ecosystem that has long hindered rapid response to global challenges like food security and disease outbreaks. By autonomously handling tasks from data discovery to model training, PPE significantly reduces the time needed to generate actionable insights, transforming weeks of manual work into minutes. This advancement not only enhances prediction accuracy across various domains but also democratizes access to high-fidelity geospatial analytics, empowering researchers and policymakers to make informed decisions swiftly.
A recent study involving over 1,000 students investigates the role of ChatGPT and critical-thinking training in enhancing student performance on university assignments. The findings reveal that students who utilized ChatGPT alongside critical-thinking exercises were able to produce more original and comprehensive responses. This indicates that AI tools, when thoughtfully integrated into educational practices, can significantly improve learning outcomes by encouraging broader thinking. The research highlights the potential for AI to work in tandem with traditional educational methods, offering a new perspective on the future of learning environments.