OpenAI Introduces New Models That Can Reason with Images

OpenAI has released two new AI models that use images as part of their reasoning process, “thinking with images.” OpenAI o3 and o4-mini “are the smartest models we’ve released to date, representing a step change in ChatGPT’s capabilities for everyone from curious users to advanced researchers,” the company says. The new entries in the “o” series also have agentic capabilities and can independently “use and combine every tool within ChatGPT, including searching the web, analyzing uploaded files and other data with Python, reasoning deeply about visual inputs, and even generating images.” Continue reading OpenAI Introduces New Models That Can Reason with Images

Anthropic Adds Deep Research, Google Integration to Claude

Anthropic has upgraded its AI assistant Claude, adding Research, an autonomous capability that integrates with Google Workspace. Claude can now search and reference content in Google Docs as well as communications in Gmail and events in Calendar. “With Research, Claude can search across both your internal work context and the web to help you make decisions and take action faster than before,” Anthropic explains, turning the model into a “true virtual collaborator” for enterprise clients. The expansion puts Anthropic into more direct competition with OpenAI and Microsoft as well as Google with Gemini in the AI productivity space. Continue reading Anthropic Adds Deep Research, Google Integration to Claude

OpenAI Reportedly Has Prototype for Its Own Social Network

OpenAI is working to build a social network that will compete against Elon Musk’s X and Meta’s Instagram, reports say. Though still in the early stages, the project is revolving around an internal prototype that is said to involve a social feed that leverages ChatGPT’s image generator. It’s unclear if an OpenAI social app would be standalone or integrated with ChatGPT, but either way it would most likely heighten the competition between rivals Musk and OpenAI CEO Sam Altman, who recently fended off an unsolicited offer by Musk to purchase his company for $97.4 billion. Continue reading OpenAI Reportedly Has Prototype for Its Own Social Network

OpenAI’s Affordable GPT-4.1 Models Place Focus on Coding

OpenAI has launched a new series of multimodal models dubbed GPT-4.1 that represent what the company says is a leap in small model performance, including longer context windows and improvements in coding and instruction following. Geared to developers and available exclusively via API (not through ChatGPT), the 4.1 series comes in three variations: in addition to the flagship GPT‑4.1, GPT‑4.1 mini and GPT‑4.1 nano, OpenAI’s first nano model. Unlike Web-connected models (which have “retrieval-augmented generation,” or RAG) and can access up-to-date information, they are static knowledge models. Continue reading OpenAI’s Affordable GPT-4.1 Models Place Focus on Coding

Netflix Tests Content Recommendations Powered by OpenAI

Netflix is testing a new recommendation engine that uses OpenAI technology to suggest viewing options based on input that goes beyond the usual parameters of cast and genre. The system is being introduced gradually and is already available in Australia and New Zealand where subscribers must opt-in to try it out, reports say, noting it allows input of more nuanced parameters, including mood, to populate search results. The partnership underscores OpenAI’s efforts to have its technology applied practically and commercially as it seeks to transition from a non-profit to a for-profit public benefit business structure. Continue reading Netflix Tests Content Recommendations Powered by OpenAI

Deep Cogito Is Out of Stealth with Hybrid Reasoning Models

San Francisco-based AI startup Deep Cogito has released five AI models in preview, making them available under an open-source license agreement. The models come in sizes 3B, 8B, 14B, 32B and 70B, with plans to release 109B, 400B and 671B versions in the weeks and months ahead. As for the current models, “each outperforms the best available open models of the same size, including counterparts from Meta, DeepSeek and Alibaba, across most standard benchmarks,” Deep Cogito claims, noting that the 70B model in particular “outperforms the newly released Llama 4 109B MoE model.” Continue reading Deep Cogito Is Out of Stealth with Hybrid Reasoning Models

Non-Profit Sentient Launches New ‘Open Deep Search’ Model

Sentient, a year-old non-profit backed by Peter Thiel’s Founders Fund, has released Open Deep Search (ODS), an open-source framework that leverages existing LLMs to enhance search and reasoning capabilities. Essentially a system of custom plugins and tools, ODS works with DeepSeek’s open-source R1 model as well as proprietary systems like OpenAI’s GPT-4o and Anthropic’s Claude to deliver advanced search functionality. That modular aspect is in fact ODS’s main innovation, its creators say, claiming it beats Perplexity and OpenAI’s GPT-4o Search Preview on benchmarks for accuracy and transparency. Continue reading Non-Profit Sentient Launches New ‘Open Deep Search’ Model

OpenAI Closes the Largest Private Tech Funding Round Ever

OpenAI has closed a $40 billion funding round, a record for a private tech firm. The infusion gives the nine-year-old San Francisco startup a $300 billion valuation making it the second most richly apprised private firm in the world, second only to SpaceX at $350 billion and tied with ByteDance, according to CNBC. The round was led by SoftBank Group contributing $30 billion, which likely gives the Japanese holding company the second largest stake, after Microsoft, which is said to have received a commitment for 49 percent of any profits in exchange for nearly $14 billion. Continue reading OpenAI Closes the Largest Private Tech Funding Round Ever

Runway Gen-4 Tackles AI’s Elusive Video Scene Consistency

Runway has introduced a new video generation model, launching a next phase of competition that could transform film production. Notably, its Gen-4 system improves the consistency of characters, locations and objects across multiple scenes, an elusive prospect for most AI video generators. The New York-based startup calls its new development “a step towards Universal Generative Models that understand the world.” The key, Runway says, is to provide a single reference image of the character, item or environment as part of the model’s project material. Runway Gen-4 can generate 5- and 10-second clips at 720p resolution. Continue reading Runway Gen-4 Tackles AI’s Elusive Video Scene Consistency

Amazon’s Nova Model Series Includes Nova Act for AI Agents

Amazon is formally rolling out its new Nova family of foundation models. Teased at the re:Invent conference hosted by AWS, details of the new multimodal series began leaking out this month. As part of the move, Amazon is diving into the agentic AI business with a new model called Nova Act, which is now in research preview. Nova Act is designed to control Web browser actions and independently tackle simple tasks. A Nova Act SDK is also being made available to allow developers to customize their own agents using the general-purpose Nova. The company is pushing for agents to help streamline business productivity. Continue reading Amazon’s Nova Model Series Includes Nova Act for AI Agents

Elon Musk Announces xAI Corporation Will Purchase X Social

Just prior to the start of the weekend, Elon Musk announced that his artificial intelligence company xAI is acquiring his social media platform X (formerly Twitter) “in an all-stock transaction,” valuing xAI at $80 billion and X at $33 billion ($45 billion less $12 billion in debt). The merger has the potential to create a powerful GenAI-powered content platform. The billionaire purchased Twitter in late 2022 for $44 billion, following months of legal skirmishes. According to Musk, X currently touts more than 600 million active users, while “xAI has rapidly become one of the leading AI labs in the world, building models and data centers at unprecedented speed and scale.” Continue reading Elon Musk Announces xAI Corporation Will Purchase X Social

Ant Group Stacks Chips to Reduce Development Costs for AI

China’s Ant Group is using local semiconductors to train AI at a cost that is 20 percent less than companies typically spend, according to reports. Ant used domestic chips — from companies including Alibaba, an investor in Ant, and Huawei — to launch a unique Mixture of Experts (MoE) training approach that produced results commensurate to training with Nvidia H800 chips. Ant is the latest Chinese company to focus on low cost training, joining a competition triggered by DeepSeek, which in January announced it could build AI comparable to the models released by U.S. companies like OpenAI, Anthropic and Google for billions less. Continue reading Ant Group Stacks Chips to Reduce Development Costs for AI

Alibaba’s Powerful Multimodal Qwen Model Is Built for Mobile

Alibaba Cloud has released Qwen2.5-Omni-7B, a new AI model the company claims is efficient enough to run on edge devices like mobile phones and laptops. Boasting a relatively light 7-billion parameter footprint, Qwen2.5-Omni-7B understands text, images, audio and video and generates real-time responses in text and natural speech. Alibaba says its combination of compact size and multimodal capabilities is “unique,” offering “the perfect foundation for developing agile, cost-effective AI agents that deliver tangible value, especially intelligent voice applications.” One example would be using a phone’s camera to help a vision impaired-person navigate their environment. Continue reading Alibaba’s Powerful Multimodal Qwen Model Is Built for Mobile

OpenAI Delivers Native GPT-4o Image Generator to ChatGPT

OpenAI has activated the multimodal image generation capabilities of GPT-4o, making it available to ChatGPT users on the Plus, Pro, Team and Free tiers. It replaces DALL-E 3 as the default image generator for the popular chatbot. GPT-4o’s accuracy with text, understanding of symbols and precision with prompts combined with well multimodal capabilities that allow the model to take cues from visual material have transformed its image capabilities from largely unpredictable to “consistent and context-aware,” resulting in “a practical tool with precision and power,” claims OpenAI. Continue reading OpenAI Delivers Native GPT-4o Image Generator to ChatGPT

Google Debuts Next-Gen Reasoning Models with Gemini 2.5

Google has released what it calls its most intelligent AI model yet, Gemini 2.5. The first 2.5 model release, an experimental version of Gemini 2.5 Pro, is a next-gen reasoning model that Google says outperformed OpenAI o3-mini and Claude 3.7 Sonnet from Anthropic on common benchmarks “by meaningful margins.” Gemini 2.5 models “are thinking models, capable of reasoning through their thoughts before responding, resulting in enhanced performance and improved accuracy,” according to Google. The new model comes just three months after Google released Gemini 2.0 with reasoning and agentic capabilities. Continue reading Google Debuts Next-Gen Reasoning Models with Gemini 2.5