Toronto, Canada
Cohere
The enterprise retrieval specialist. Cohere focuses on retrieval-augmented generation and tool-calling rather than topping leaderboards. Command R+ is built for citation-accurate pipelines, with open weights so you're never locked in.
Models
Recent news
Articles mentioning Cohere models
NVIDIA's AI Agents Now Understand Long-Term Context
NVIDIA has introduced a major upgrade to its AI agents, enabling them to maintain and understand long-term context across extended interactions. This breakthrough allows the agents to track changes in messages, decisions, projects, and obligations over time-something they couldn't do before. This advancement is significant because it bridges a critical gap in AI capabilities, where most systems struggle to retain information across conversations. For developers and researchers, this means more reliable and coherent interactions with AI tools, as the agents can now handle evolving tasks without losing track of previous exchanges. Look for more applications of context-aware AI in customer service, project management, and complex problem-solving in the coming months.
NVIDIA Dev Blog4w ago
AI-Powered Solutions Transform Healthcare and Root-Cause Analysis
Cohere Health has developed a new system using Amazon Bedrock AgentCore to digitize clinical policies, which are crucial for prior authorization in healthcare. This process was previously manual and time-consuming, but now it can be automated by converting static documents into structured data. This innovation helps health plans reduce bottlenecks and improve efficiency while maintaining clinical oversight. Meanwhile, researchers at the TReNDS Center have created an AI-driven solution to automatically analyze and resolve technical errors. By integrating Amazon Bedrock with CloudWatch and Lambda, they enable real-time error detection and root-cause analysis. This reduces the time engineers spend investigating issues, from 15-30 minutes for simple errors to hours for complex ones. Amazon Bedrock AgentCore also introduces temporal policies to secure AI agents by evaluating their session history. This ensures that actions are authorized based on context, preventing potential misuse like unauthorized transactions or data manipulation. As these technologies evolve, they promise to enhance both healthcare operations and software reliability across industries.
AWS ML Blog1mo ago
AI Tools Are Transforming Technical Research - But Not Always for the Better
AI tools are rapidly changing how technical research is conducted, with both benefits and drawbacks. Recent advancements like Claude Code have enabled AI agents to perform complex coding tasks, run experiments, and even write up research findings, making researchers more efficient. However, this shift has also led to challenges in peer review, as some submissions appear to be low-quality or nonsensical, likely generated by AI without proper oversight. At the Mechanistic Interpretability Workshop, organizers noticed a significant increase in submissions that seemed to resemble "AI slop"-content that appears coherent but lacks depth. Reviewers found it difficult to assess these papers, often spending extra time trying to understand abstracts that didn't clearly state their contributions. To address this issue, workshop chairs used Pangram, an AI-text detector, to analyze submissions and reviews, revealing the extent of AI-generated content. Looking ahead, researchers need to find a balance between leveraging AI's capabilities and maintaining the quality and rigor of academic work. As AI tools become more advanced, it will be crucial to develop guidelines and detection methods to ensure that research remains meaningful and credible.
LessWrong2mo ago
DeepSeek Boosts AI Generation Speed with DSpark Module
DeepSeek has introduced a new module called DSpark, which significantly enhances the speed of large language model (LLM) generation without compromising on quality. By using speculative decoding, DSpark addresses two major issues in AI production-draft quality and computational waste. This technique allows models to generate text more efficiently, boosting per-user generation speed by 60 to 85 percent. The innovation matters because it directly impacts both developers and end-users. For developers, integrating DSpark can lead to faster and more resource-efficient AI systems. For users, this means receiving responses quicker without sacrificing the accuracy or coherence of the generated content. The technology could also reduce costs for companies relying on AI-powered services. Looking ahead, DeepSeek plans to expand its application beyond LLMs, potentially revolutionizing other areas of AI development and deployment.
Analytics Vidhya2mo ago
AI Agents Now Remember Conversations Across Days
AI agents can now remember details from previous interactions across days, marking a significant leap in their capabilities. Previously limited to handling single questions or short exchanges, these advancements allow agents like NVIDIA's to maintain context over extended periods. For instance, an agent can recall past user preferences and tailor responses accordingly. This improvement is crucial for developers and researchers aiming to create more intuitive AI systems that match human-like interaction. By retaining information across multiple exchanges, these agents can provide more coherent and personalized assistance, enhancing user experience in applications like customer support or personal assistants. Looking ahead, expect further developments in memory retention and contextual understanding, potentially enabling even longer-term recall and more sophisticated conversational flows.
NVIDIA Dev Blog3mo ago