Fields
Multimodal
Models and embeddings that combine text, images, speech and video
Stories: 2 · newest first
October 7, 2026
-
OpenAI launches Decisions API on decision model gpt-6-luna: $0.10 per 1M input tokens, output free (community report)
According to Simon Willison, OpenAI has launched the 'Decisions API' it previewed at DevDay last week. The API works the same way as Jev...
-
Twelve Labs launches Pegasus 1.6, a VLM that turns first-person video into robot training data (report)
Twelve Labs said it has launched Pegasus 1.6, a vision-language model specialized for collecting physical AI training data. It...