Categories

Artificial Intelligence

Gemini 2.5: context scaling and multimodality

Google publicó Gemini 2.5 Pro en vista previa en marzo y la versión general llegó en junio. El salto respecto a Gemini 2.0 no está solo en puntuaciones sino en dos frentes prácticos: ventana de contexto utilizable en serio y multimodalidad que deja de ser demostración para convertirse en herramienta.

Artificial Intelligence

GPT-4o: OpenAI’s Native Multimodality

GPT-4o is the OpenAI model presented on May 13, 2024, that fuses text, image, and audio into a single native model, without separate pipelines. It delivers roughly 320-millisecond conversational latency, better multimodal understanding, and a price 50% lower than GPT-4 Turbo.