Terminal-Bench 2.0 and Harbor: New AI Testing Tools
Terminal-Bench 2.0 and Harbor offer improved tools for testing AI agents, enhancing consistency and scalability in containerized environments.
Terminal-Bench 2.0 and Harbor offer improved tools for testing AI agents, enhancing consistency and scalability in containerized environments.
Seven families have filed lawsuits against OpenAI, claiming ChatGPT’s role in suicides and delusions due to insufficient safeguards in the GPT-4o model.
OpenAI’s economic research team faces internal conflict over publishing practices, raising concerns about AI advocacy and transparency.
The Linux Foundation’s Agentic AI Foundation seeks to standardize AI agent interoperability, supported by OpenAI, Anthropic, and Block.
OpenAI, Anthropic, and Block have founded the Agentic AI Foundation to promote open standards for AI agents, enhancing interoperability and innovation.
The European Commission is investigating Google’s AI search tools for potential breaches of competition laws, focusing on content use and market fairness.
Microsoft is set to invest $17.5 billion in India by 2029, enhancing its AI and cloud infrastructure, and expanding data centers and skilling programs.
India’s new proposal could reshape AI operations by requiring royalties for using copyrighted content in model training.
Anthropic and Accenture have signed a three-year strategic partnership to enhance AI capabilities in the enterprise sector.
Mistral AI has launched Devstral 2, a new AI model for coding, to compete with major AI labs like Anthropic.