AI Europe
  • Europe
  • Europa
  • Britain
  • France
  • Germany
  • Italy
  • Spain
  • Japan
  • Canada
  • Africa
  • People
  • AI
  • …
    • Afrique
    • Netherlands
    • Poland
  • Agentic AI
  • AGI
  • AI
  • Anthropic
  • Google
  • Microsoft
  • OpenAI
  • xAI
AI Europe
  • Europe
  • Europa
  • Britain
  • France
  • Germany
  • Italy
  • Spain
  • Japan
  • Canada
  • Africa
  • People
  • AI
  • …
    • Afrique
    • Netherlands
    • Poland

Browsing Tag

benchmarking

8 posts
GGoogle
Google expands AI infrastructure with Lustre & C4N
Read More

Google expands AI infrastructure with Lustre & C4N

  • August 3, 2026
Google Cloud has moved its Managed Lustre storage service into general availability and has also made its C4N…
AAgentic AI
Google Cloud tests AI agents with ambiguity-based benchmarks
Read More

Google Cloud tests AI agents with ambiguity-based benchmarks

  • July 13, 2026
Google Cloud has outlined a way to evaluate AI agents by varying the ambiguity of test queries instead…
AAgentic AI
Copilot appears against a decorative background with scattered green squares.
Read More

Evaluating performance and efficiency of the GitHub Copilot agentic harness across models and tasks

  • June 25, 2026
While the model provides the raw intelligence, the harness shapes how effectively that intelligence is applied. The GitHub Copilot agentic…
AAI
GAM Launches A Chatbot For Troubleshooting Ad Campaigns
Read More

GAM Launches A Chatbot For Troubleshooting Ad Campaigns

  • June 18, 2026
Google Ad Manager (GAM) wants to make it easier for publisher ad ops teams to troubleshoot campaigns. And…
OOpenAI
Omnea Brings Procurement Data to Claude, Cohere North, and ChatGPT with Industry-first MCP
Read More

New AI Benchmarking Reveals Leading AI Chatbots––Including Claude, ChatGPT, and Gemini—Avoid Harm, but Still Need More Support for High Risk Conversations

  • May 13, 2026
mpathic launches mPACT, a clinician-led benchmark for evaluating how AI models perform in high-risk scenarios, including suicide risk,…
AAnthropic
Causal Dynamics Lab launches Cielara Code for AI coding
Read More

Causal Dynamics Lab launches Cielara Code for AI coding

  • May 5, 2026
Causal Dynamics Lab has launched Cielara Code, a tool it says outperformed Anthropic’s Claude Code and OpenAI’s Codex…
MMicrosoft
Microsoft to test third-party AI models for incorporation in its security offerings
Read More

Microsoft to test third-party AI models for incorporation in its security offerings

  • April 22, 2026
Microsoft will be evaluating third-party artificial intelligence systems to pair with its network monitoring to see if powerful…
MMicrosoft
Microsoft to test third-party AI models for incorporation in its security offerings
Read More

Microsoft to test third-party AI models for incorporation in its security offerings

  • April 22, 2026
Microsoft will be evaluating third-party artificial intelligence systems to pair with its network monitoring to see if powerful…
AI Europe
www.europesays.com