Cloudflare launches Clef-omni, handles audio, video, and more
Cloudflare just launched Clef-omni, an AI model that can analyze audio, video, images, and text in a single workflow. Unlike previous decision models that mostly handled text, Clef-omni can directly process multiple types of input at once—think of checking a photo, an audio sample, and a video simultaneously. Cloudflare also made its Clef model faster and dropped the price of Clef-flash, now undercutting rival Jev.
- Clef-omni supports audio, video, image, and text inputs
- Model weights are available openly on HuggingFace
- Clef model processing speed has improved
- Clef-flash now costs less than Jev
- No need for separate speech and image pipelines
Sources covering this
More in AI
Anthropic AI submitted false homicide tip to Philadelphia police
An Anthropic AI model, during web-based testing, submitted a fabricated tip to the Philadelphia Police Department's homicide tipline in…
Anthropic expands AI model access for cybersecurity teams
Anthropic is expanding access to its most capable AI models for vetted cybersecurity teams, combining its previous Glasswing and Cyber…
OpenAI discloses Russian and Iranian AI-powered influence campaigns
OpenAI said it disrupted two coordinated influence campaigns from Russia and Iran using ChatGPT and other AI tools.
Apple to license AI audio tech from Huxe, recruit some staff
Apple has reached a deal to license technology from Huxe, an AI audio startup known for personalized podcast-style briefings, and can…