Skip to main content

Share story

Science

Preprint introduces Arabic-language LLM safety benchmark

A preprint introduces the Arabic Safety Index, a human-curated red-teaming benchmark with 801 prompts across eight safety categories and eight attack strategies. Its authors evaluated seven Arabic-capable models and report that most failed to defend against 50% of unsafe prompts. The study also reports that direct and obfuscation-based attacks were most effective and that automated safety judges performed poorly against human annotators.
Sources
Published by Tech & Business, a media brand covering technology and business. This story was sourced from cs.AI updates on arXiv.org, cs.AI updates on arXiv.org and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
Infrastructure Capital
Infrastructure Capital

Cloudflare acquires Deno, a rival in developer infrastructure

Cloudflare is buying Deno, the startup co-founded by Node.js creator Ryan Dahl, The New Stack reported. Deno had developed an open-source alternative to Cloudflare Workers, making the acquisition a purchase of a longtime competito...

AI Infrastructure
AI Infrastructure

Synopsys explores Chinese AI partnerships for chip design tools

Synopsys says it is exploring partnerships with Chinese AI labs to develop AI-powered chip design tools for the Chinese market, Nikkei Asia reports. The chip design software company wants to help Chinese companies develop chips fa...

Capital Products
Capital Products

Firmus IPO collapses after investors reject $30B valuation

Firmus' IPO collapsed in 48 hours after US fund managers judged its $30B valuation too high, Bloomberg reported, citing sources. The Nvidia-backed company recorded $51M in FY 2026 revenue, the financial comparison at the center of...