All posts tagged: LLM

Vercel Releases v0 AI Model for Web Application Development, Compatible with OpenAI API

Vercel Releases v0 AI Model for Web Application Development, Compatible with OpenAI API

Vercel, the company behind the vibe coding platform for web application development, v0, is now releasing an artificial intelligence (AI) model. Announced on Thursday, the v0 AI model is available via an application programming interface (API), as well as other formats. It is the first AI model developed by the company. The San Francisco-based cloud platform as a service (PaaS) company says the AI model is specialised for web application development (front and back-end) tasks, and is compatible with OpenAI’s API. Vercel’s v0 AI Model Can Develop Websites and Web Apps In a post on X (formerly known as Twitter), the official handle of v0, announced the release of the new AI model. It is the same model that powers the vibe coding platform v0, and the company says it specialises in website development knowledge. It is currently available in beta via the company API, as an AI software development kit (SDK), or via AI Playground. Vercel says the AI model, officially named v0-1.0-md, supports both text and images as input and is designed to …

Anthropic Releases Claude 4 Series AI Models With Improved Coding Capability and Tool Use

Anthropic Releases Claude 4 Series AI Models With Improved Coding Capability and Tool Use

Anthropic introduced Claude 4 artificial intelligence (AI) models at its inaugural developer conference on Thursday. The San Francisco-based AI firm unveiled Claude Opus 4 and Claude Sonnet 4 models, and announced new capabilities including Extended Thinking with tool use. Opus 4 is said to be state-of-the-art (SOTA) in coding, tool use, and writing. Additionally, Claude Code is now generally available, and individuals can find its beta extensions in VS Code and JetBrains. It is also among the models available on GitHub. Anthropic Unveils Claude 4 AI Models In a newsroom post, the AI firm detailed the new models as well as the new features it is rolling out across its chatbot and application programming interface (API). Anthropic’s latest large language models (LLMs) put a heavy focus on coding capabilities and agentic functions. Both Opus 4 and Sonnet 4 are hybrid models with two modes: near-instant responses and Extended Thinking for deeper reasoning. Opus 4 is the company’s flagship-tier AI model. Calling it “the best coding model in the world,” Anthropic claimed that it scored 72.5 …

Google Unveils Video-Generation Model Veo 3 and Image Generation Model Imagen 4 at Google I/O 2025

Google Unveils Video-Generation Model Veo 3 and Image Generation Model Imagen 4 at Google I/O 2025

Google unveiled the next generation of its image and video generation artificial intelligence (AI) models on Tuesday at the I/O 2025 event. Dubbed Imagen 4 and Veo 3, these multimodal AI models arrive with new capabilities and upgrades over their predecessors. While Imagen 4 features faster generation times and improved text rendering, Veo 3 gets native audio generation capability and can integrate background sound and dialogues in generated videos. Alongside the new models, the tech giant also unveiled a new AI-powered filmmaking app dubbed Flow. What’s New With Imagen 4 and Veo 3? In a blog post, the Mountain View-based tech giant detailed the new image and video generation AI models. Imagen 4 comes almost a year after its predecessor was released. In December 2024, Google also released Veo 2 and updated Imagen 3 with new capabilities. Now, with Imagen 4, the company is focusing on generation speed and accuracy of the model. Similar to the previous generation, the latest Imagen model also supports text and images as input. The generated images witness an improvement in …

Apple is trying to get ‘LLM Siri’ back on track

Apple is trying to get ‘LLM Siri’ back on track

Apple Intelligence has been a wreck since its first features rolled out last year, and a big new report from Bloomberg’s Mark Gurman details why — and how Apple is trying to piece things back together. And much of its effort hinges on rebuilding Siri from the ground up. Gurman has reported in the past that Apple is working on what it’s internally calling ‘LLM Siri’ — a reworked, generative AI version of the company’s digital assistant. Apple’s previous approach of merging the assistant with the existing Siri hasn’t been working. Gurman describes in great detail a number of reasons why, but here’s a quick summary: Now the company is trying to rejigger its approach. Part of that is a total overhaul of Siri, rather than just trying to make generative AI work in concert with the old Siri. According to Gurman, Apple has its AI team in Zurich working on a new architecture that will “entirely build on an LLM-based engine.” Gurman reported in November last year that the company was working on this, …

Windsurf Releases SWE-1 Series AI Models Capable of Full-Process Software Development

Windsurf Releases SWE-1 Series AI Models Capable of Full-Process Software Development

Windsurf, an artificial intelligence (AI) no-code or ‘vibe coding’ platform, released a series of AI models on Thursday. Dubbed SWE-1, these models are focused on complex software engineering tasks, and not just writing code. The series comprises SWE-1, SWE-1-lite, and SWE-1-mini, where each model is designed for specific use cases. The lite and mini versions are available to all Windsurf users, whereas the frontier SWE-1 model is only available to subscribers. The company has yet to reveal pricing for and availability for the frontier coding model. Windsurf Unveils New Software Engineering AI Models The California-based AI firm detailed the new series of AI models in a blog post. SWE-1 is designed to tackle the complex tasks and responsibilities of a human, and not just write and edit code. The company says that while coding models have improved, their scope has not increased significantly. Most models are still trained to produce code that compiles and passes unit tests, but that represents only a small part of what software engineers do. The company added that the next …

Meta Delays Release of Its ‘Behemoth’ AI Model: Report

Meta Delays Release of Its ‘Behemoth’ AI Model: Report

Meta Platforms is delaying the release of its flagship “Behemoth” AI model due to concerns about its capabilities, the Wall Street Journal reported on Thursday, citing people familiar with the matter. Company engineers are struggling to significantly improve the capabilities of its Behemoth large-language model, resulting in staff questions about whether improvements over earlier versions are significant enough to justify public release, the report said. Early in its development, Behemoth was internally scheduled for release in April to coincide with Meta’s inaugural AI conference for developers, but later pushed an internal target for the model’s launch to June, according to the report. It has now been delayed to fall or later, the report said. The social media giant did not immediately respond to a Reuters request for comment. Meta had said in April it was previewing Llama 4 Behemoth, which it called “one of the smartest LLMs in the world and our most powerful yet to serve as a teacher for our new models”. It released the latest version of its LLM Llama, called the …

Xiaomi MiMo AI Models Launched With Efficient Reasoning, Small Size

Xiaomi MiMo AI Models Launched With Efficient Reasoning, Small Size

Xiaomi on Tuesday released an open-source reasoning-focused artificial intelligence (AI) model. Dubbed MiMo, the family of reasoning models innovate the optimisation of reasoning capability in a relatively smaller parameter size. This is also the first open-source reasoning model by the tech giant, and it competes with Chinese models such as DeepSeek R1 and Alibaba’s Qwen QwQ-32B, and global reasoning models including OpenAI’s o1 and Google’s Gemini 2.0 Flash Thinking. The MiMo family comprises four different models, each with unique use cases. Xiaomi’s MiMo Reasoning AI Model to Compete With DeepSeek R1 With the MiMo series of AI models, Xiaomi researchers aimed to solve the size problem in reasoning AI models. Reasoning models (at least ones that can be measured) have around 24 billion or more parameters. The large size is kept to achieve uniform and simultaneous improvements in both coding and mathematical capabilities of large language models, something considered difficult to achieve with smaller models. In comparison, MiMo features seven billion parameters, and Xiaomi claims that its performance matches OpenAI’s o1-mini and outperforms several reasoning …

Microsoft Releases Largest 1-Bit LLM, Letting Powerful AI Run on Some Older Hardware

Microsoft Releases Largest 1-Bit LLM, Letting Powerful AI Run on Some Older Hardware

Microsoft researchers claim to have developed the first 1-bit large language model with 2 billion parameters. The model, BitNet b1.58 2B4T, can run on commercial CPUs such as Apple’s M2. “Trained on a corpus of 4 trillion tokens, this model demonstrates how native 1-bit LLMs can achieve performance comparable to leading open-weight, full-precision models of similar size, while offering substantial advantages in computational efficiency (memory, energy, latency),” Microsoft wrote in the project’s Hugging Face depository. What makes a bitnet model different? Bitnets, or 1-bit LLMs, are compressed versions of large language models. The original 2-billion parameter scale model trained on a corpus of 4 billion tokens was shrunken down into a version with drastically reduced memory requirements. All weights are expressed as one of three values: -1, 0, and 1. Other LLMs might use 32-bit or 16-bit floating-point formats. SEE: Threat actors can inject malicious packages into AI models that resurface during “vibe coding.” In the research paper, which was posted on Arxiv as a work in progress, the researchers detail how they created the …

Google Releases Cost-Efficient and Low-Latency Gemini 2.5 Flash AI Model

Google Releases Cost-Efficient and Low-Latency Gemini 2.5 Flash AI Model

Google released its second artificial intelligence (AI) model in the Gemini 2.5 family on Thursday. Dubbed Gemini 2.5 Flash, it is a cost-efficient low-latency model which is designed for tasks requiring real-time inference, conversations at scale, and those which are generalistic in nature. The Mountain View-based tech giant will soon make the AI model available on both the Google AI Studio as well as Vertex AI to help users and developers access the Gemini 2.5 Flash, and build applications and agents using it. Gemini 2.5 Flash Is Now Available on Vertex AI In a blog post, the tech giant detailed its latest large language model (LLM). Alongside announcing the debut of the Flash model, the post also confirmed that the Gemini 2.5 Pro model is now available on Vertex AI. Differentiating between the use cases of the two models, Google said the Pro model is ideal for tasks that require intricate knowledge, multi-step analyses, and making nuanced decisions. On the other hand, the Flash model prioritises speed, low latency, and cost efficiency. Calling it a …

Meta Unveils Llama 4 AI Series Featuring New Expert-Based Architecture

Meta Unveils Llama 4 AI Series Featuring New Expert-Based Architecture

Image: Meta Meta unveiled on April 5 its new AI model series: Llama 4, which includes Llama 4 Maverick and Llama 4 Scout, tailored for conversation and processing large files, respectively, along with an unreleased “teacher” model called Llama 4 Behemoth. Llama 4 is Meta’s first series to adopt a “mixture of experts (MoE) architecture.” This approach activates only select parts of the neural network, referred to as the “experts,” to handle specific subtasks. The task will be broken down into subtasks and each routed to the most appropriate experts, improving resource efficiency. What are the specifics about Llama 4 Maverick and Scout? Llama 4 Maverick features 128 experts and 17 billion active parameters, which represent the portion of a model’s knowledge used to process a given input. Meta describes it as the “product workhorse model for general assistant and chat use cases,” specialising in image interpretation and creative writing. Interestingly, Mark Zuckerberg’s company boasts that Maverick offers “a best-in-class performance to cost ratio” when it comes to conversations. Cost has been playing on the …