❌

Normal view

Google's New LiteRT Accelerator Supercharges AI Workloads on Snapdragon-powered Android Devices

1 December 2025 at 04:00

Google has introduced a new accelerator for LiteRT, called Qualcomm AI Engine Direct (QNN), to enhance on-device AI performance on Qualcomm-powered Android devices equipped with Snapdragon 8 SoCs. The accelerator delivers significant gains, offering up to a 100x speedup over CPU execution and 10x over GPU.

By Sergio De Simone

Android GenAI Prompt API Enables Natural Language Requests with Gemini Nano

7 November 2025 at 03:00

The ML Kit GenAI Prompt API, now available in alpha, enables Android developers to send natural language and multimodal requests to Gemini Nano running on-device, extending the text summarization and image description capabilities introduced with the initial GenAI release.

By Sergio De Simone
  • βœ‡InfoQ
  • OpenAI Study Investigates the Causes of LLM Hallucinations and Potential Solutions Sergio De Simone
    In a recent research paper, OpenAI suggested that the tendency of LLMs to hallucinate stems from the way standard training and evaluation methods reward guessing over acknowledging uncertainty. According to the study, this insight could pave the way for new techniques to reduce hallucinations and build more trustworthy AI systems, but not all agree on what hallucinations are in the first place. By Sergio De Simone
     

OpenAI Study Investigates the Causes of LLM Hallucinations and Potential Solutions

13 October 2025 at 00:00

In a recent research paper, OpenAI suggested that the tendency of LLMs to hallucinate stems from the way standard training and evaluation methods reward guessing over acknowledging uncertainty. According to the study, this insight could pave the way for new techniques to reduce hallucinations and build more trustworthy AI systems, but not all agree on what hallucinations are in the first place.

By Sergio De Simone

Google Introduces VaultGemma: An Experimental Differentially Private LLM

26 September 2025 at 02:00

VaultGemma is a 1B-parameter Gemma 2-based LLM that Google trained from scratch using differential privacy (DP) with the aim of preventing the model from memorizing and later regurgitating training data. While still a research model, VaultGemma could enable applications cases in healthcare, finance, legal, and other regulated sectors.

By Sergio De Simone

Anthropic Investigates How Large Language Models Develop a Character

12 August 2025 at 20:00

Recent research by Anthropic engineers explores identifiable patterns of activity that seems to give rise to an emerging personality. These traits, known as persona vectors, help explain how a model's personality shifts over its lifecycle and lay the groundwork for better controlling those changes.

By Sergio De Simone
❌