Artificial intelligence
-
Cohere AI Releases Cohere Transcribe: SOTA Automatic Speech Recognition (ASR) Model Powering Enterprise Speech Intelligence
In the case of enterprise AI, the bridge between unstructured audio and physical text is often a bottleneck of proprietary…
Read More » -
Tencent AI Open Sources Covo-Audio: 7B Speech Language Model and Suggestive Line for Real-Time Audio Conversations and Consultations
Tencent AI Lab has been released Covo-Audioparameter 7B-end-to-end Large Audio Language Model (LALM). The model is designed to integrate speech…
Read More » -
AI system learns to keep shop robot traffic efficient | MIT News
Inside a large private warehouse, hundreds of robots race through the aisles as they collect and distribute items to fill…
Read More » -
How to Build a Vision-Driven Web Agent with MolmoWeb-4B Using Multimodal Reasoning and Action Prediction
def parse_click_coords(action_str): """ Extract normalised (x, y) coordinates from a click action string. e.g., 'click(0.45, 0.32)' -> (0.45, 0.32) Returns…
Read More » -
Expanding citizen science with computer vision for fish monitoring | MIT News
Each spring, herring migrate from Massachusetts’ coastal waters to begin their annual journey up rivers and streams to freshwater. River…
Read More » -
The Wristband enables wearers to control the robot’s hand with their movements | MIT News
The next time you scroll through your phone, take a moment to appreciate this experience: This seemingly unusual action is…
Read More » -
NVIDIA AI Introduces PivotRL: A New AI Framework That Achieves Higher Agent Accuracy with 4x Fewer Outputs and More Efficient Turns
After training Large-scale Language Modelers (LLMs) for long-horizon agent tasks—such as software engineering, web browsing, and the use of complex…
Read More » -
Google Introduces TurboQuant: A New Compression Algorithm That Reduces LLM Key Value Cache Memory by 6x and Delivers Up to 8x Speedup, All with Zero Loss of Accuracy
The scaling of large-scale language models (LLMs) is increasingly constrained by the memory interface between High-Bandwidth Memory (HBM) and SRAM.…
Read More » -
Paged Attention to Major Language Models LLMs
When using LLMs at scale, the real limitation is GPU memory rather than computation, mainly because each application needs a…
Read More » -
This AI Paper Introduces TinyLoRA, a 13-Parameter Fine-Tuning Method That Achieves 91.8 Percent of GSM8K on Qwen2.5-7B
Researchers from FAIR on the Meta, Cornell Universityagain Carnegie Mellon University showed that large-scale linguistic models (LLMs) can learn reasoning…
Read More »