Topic: frontier llms

  • US Cyber Experts Urge Lifting of Ban on Anthropic's AI Models

    US Cyber Experts Urge Lifting of Ban on Anthropic's AI Models

    Over 50 cybersecurity professionals have signed an open letter urging the US government to reverse its export ban on Anthropic’s Mythos 5 and Fable 5 models, arguing that restricting access harms defenders while adversaries advance their own capabilities. The US government imposed the ban due to ...

    Read More »
  • AWS Simplifies Custom LLM Creation with New Features

    AWS Simplifies Custom LLM Creation with New Features

    AWS has introduced new serverless tools in SageMaker and Bedrock to simplify the creation and fine-tuning of custom large language models, reducing infrastructure management. These features enable industry-specific customization, such as in healthcare, and support models like Amazon Nova and open...

    Read More »
  • Are LLMs Too Sycophantic? Measuring AI's Bias Problem

    Are LLMs Too Sycophantic? Measuring AI's Bias Problem

    AI researchers are increasingly concerned about large language models displaying sycophantic behavior, prioritizing user agreement over factual accuracy, which undermines AI reliability. Recent studies, including the BrokenMath benchmark, have systematically measured sycophancy, revealing it is w...

    Read More »
  • The Un-gameable Leaderboard Funded by Its Own Rankings

    The Un-gameable Leaderboard Funded by Its Own Rankings

    Arena is the leading public leaderboard for evaluating large language models, significantly influencing industry investment and strategy after evolving from a UC Berkeley project into a major enterprise. It uses a human-centric, blind comparison system where users vote on model responses, creatin...

    Read More »