Search results for Llama 3.3
AI & Enterprise
Elon Musk scores 992, Tim Cook 986 in service scoring AI recognition of famous names
A web service called In The Weights has been launched to compare how strongly different generative AI models have learned specific people’s names. It assigns a “strength score” by asking each model who a person is and aggregating up to 10 candidate results with brief descriptions and confidence levels. Examples showed Elon Musk at 992 and Apple CEO Tim Cook at 986. The service also flags possible hallucinations and misidentifications.
AI & Enterprise
GitHub tool bypasses AI safeguards in some Meta, Google open-weight models, test finds
Tests found that safeguards in some open-weight AI models released by Meta and Google could be disabled within minutes using a tool posted on GitHub. After safety controls were removed, Meta\'s Llama 3.3 and Google\'s Gemma 3 responded to risky questions they should have refused. The tool\'s creator said it was used to make more than 3,500 modified models, with cumulative downloads exceeding 13 million.
AI & Enterprise
Hiring AI shows bias, gives edge to resumes written by same model
A study found that AI used in recruitment tends to rate resumes more highly when they are written by the same chatbot model. Researchers from the University of Maryland, the National University of Singapore and Ohio State University called it “AI self-preference bias” and tested it using 2,245 human-written resumes from LiveCareer.com. Across multiple models, evaluators often favored same-model AI summaries even after controlling for writing quality. Mitigation tests reduced the bias.