Multimodal AI

Auto Added by WPeMatico

Banner for the AI & Big Data Expo event series.

Guardoc Health processes clinical documentation using Amazon Nova models

Guardoc Health says it processes over one million clinical documents daily using Amazon Nova models through Bedrock. Bringing AI into clinical documentation comes down to a specific kind of risk calculation. Get it wrong and the errors compound into denied Medicare claims under the Patient-Driven Payment Model, audit fines, litigation exposure, and in the worst […]

Guardoc Health processes clinical documentation using Amazon Nova models Read More »

Banner for the AI & Big Data Expo event series.

Google’s Gemini 3.6 Flash targets enterprise agent token costs

Google has released Gemini 3.6 Flash and 3.5 Flash-Lite as new workhorses designed to cut latency and token costs for enterprise AI agents. The economics of running autonomous software agents inside a production environment come down to a fixed equation few vendors advertise directly. A model needs to reason through a multi-step task competently, but

Google’s Gemini 3.6 Flash targets enterprise agent token costs Read More »

Multimodal Data for Humanoid Robots

Multimodal Data for Humanoid Robots: Vision, Language, Action, Telemetry, and Context

Ask a humanoid robot to “pick up the red mug on the left and place it in the sink,” and a remarkable amount has to happen at once. The robot must see the mug, parse the instruction, plan a motion, feel the grip pressure, and understand that “the sink” is the wet basin three feet

Multimodal Data for Humanoid Robots: Vision, Language, Action, Telemetry, and Context Read More »

Banner for the AI & Big Data Expo event series.

Deploying retail AI to scale personalisation and customer insight

Optimising retail AI infrastructure drives the successful deployment of personalisation systems and real-time customer insight. Leaders are replacing static customer interaction patterns with data pipelines capable of modifying the user environment during a live session. Static layouts and broad segmentation rules fail to satisfy modern conversion targets. Deployments demonstrate that traditional demographic categorisation generates insufficient

Deploying retail AI to scale personalisation and customer insight Read More »

Banner for the AI & Big Data Expo event series.

Omio scales travel product development using OpenAI models

Omio integrates OpenAI models across its engineering operations to accelerate travel product development and launch booking interfaces. The multimodal travel platform coordinates operations with over 3,000 transportation providers across 47 countries. Omio explicitly rejects the superficial addition of technology to outdated internal processes. The company’s CTO, Tomas Vocetka, requires all internal functions to completely redesign

Omio scales travel product development using OpenAI models Read More »

Physical AI raises governance questions for autonomous systems

Governance around Physical AI is becoming harder as autonomous AI systems move into robots, sensors, and industrial equipment. The issue is not only whether AI agents can complete tasks. It is how their actions are tested, monitored, and stopped when they interact with real-world systems. Industrial robotics already provides a large base for that discussion.

Physical AI raises governance questions for autonomous systems Read More »

NVIDIA and Google infrastructure cuts AI inference costs

At the Google Cloud Next conference, Google and NVIDIA outlined their hardware roadmap designed to address the cost of AI inference at scale. The companies detailed the new A5X bare-metal instances, which run on NVIDIA Vera Rubin NVL72 rack-scale systems. Through hardware and software codesign, this architecture aims to deliver up to ten times lower

NVIDIA and Google infrastructure cuts AI inference costs Read More »

Citizen developers now have their own Wingman

A vibe-coding application creation company, Emergent, has released Wingman, an autonomous agent that can address and take control of the applications used to manage daily tasks. The company’s press release states: “The best technology should be accessible to everyone”, and cites the difficulty that users without a technical background have in creating software applications. It

Citizen developers now have their own Wingman Read More »

Meta has a competitive AI model but loses its open-source identity

The open-source AI movement has never lacked for options. Mistral, Falcon, and a growing field of open-weight models have been available to developers for years. But when Meta threw its weight behind Llama, something shifted. A company with three billion users, vast compute resources, and the credibility of a tech giant was now building openly,

Meta has a competitive AI model but loses its open-source identity Read More »

Automating complex finance workflows with multimodal AI

Finance leaders are automating their complex workflows by actively adopting powerful new multimodal AI frameworks. Extracting text from unstructured documents presents a frequent headache for developers. Historically, standard optical character recognition systems failed to accurately digitise complex layouts, frequently converting multi-column files, pictures, and layered datasets into an unreadable mess of plain text. The varied

Automating complex finance workflows with multimodal AI Read More »