TECH & OTHER NEWS

Microsoft may have an audio-to-image generator in the works, new patent shows

October 15, 2024

Waveform concept — spfdigital/Getty Images

There are currently many artificial intelligence (AI) tools on the market that can take users’ text and images and transform them into images and videos that match the initial prompt. A new patent reveals that audio may soon be an input option to bring your visions to real life.

As spotted by MSPowerUser, the US Patent and Trademark Office (USPTO) posted a 20-page document filed by Microsoft on April 5, 2023, and published on October 10, 2024, that details a new AI-supported system that converts live audio into images.

Also: Adobe’s free AI video generator is here – how to try it out

This system would take an audio live stream, such as that from a meeting or lecture, and convert it into a live text transcript. The transcript would then be summarized by a large language model (LLM) and fed into a text-to-image model, where an image would be generated and output on the screen, as seen in the image below.

This system would continue to do this during the audio stream, continuously generating live images. According to Microsoft, displaying images in real-time can help make communication more effective, with visual aids keeping people more engaged and making concepts easier to understand.

“Displaying images related to verbally communicated information can enhance the effectiveness of communication by making it more engaging, memorable, and easier to understand,” said Microsoft.

Also: The best AI chatbots of 2024: ChatGPT, Copilot, and worthy alternatives

If you’re wondering whether the feature will launch soon, the answer is most likely no. Filing a patent is a long journey between producing a product or feature, and many patents never make it into the production phase and remain an idea.

However, if Microsoft does decide to launch this feature, it would likely live in Microsoft Teams, its video conferencing meeting platform, and be accessible through its AI add-on, Copilot, such as Copilot Pro or Microsoft 365 Copilot for businesses.

Artificial Intelligence

Source Link

Microsoft may have an audio-to-image generator in the works, new patent shows

Artificial Intelligence

LEAVE A REPLY Cancel reply

TECH NEWS

Everything Old is New Again: AI-Driven Development and Open Source

Gen AI in Healthcare: The State of Affairs in India

Gartner Predicts Legal, Risk and Compliance Functions to Double Technology Spend...

Microsoft to End Support for Windows Mail, Calendar and People Apps...

IDC Predicts: Asia/Pacific Business Leaders to Demand 80% Success Rate on...

The Cooling Conundrum: AI and Automation Push Data Centers Toward 3X...

TOP STORIES

Seventy Percent of Economies Are Underprepared for AI Disruption

New study shows almost half of tech professionals in India believe...

Organizations Combining Organizational Learning and AI-Specific Learning Are up to 80%...

Nvidia’s AI-driven triumph over Intel powered by strategic innovations

Most banks and insurers adopt cloud solutions with the primary objective...

India’s Web3 Ecosystem Has Over 400 Firms, Karnataka Emerges as Industry...

Cyber Security

AI and Gen AI are set to transform cybersecurity for most...

ThreatQuotient Publishes 2024 Evolution of Cybersecurity Automation Adoption Research Report

Kaspersky predicts quantum-proof ransomware and advancements in mobile financial cyberthreats in...

Rising concerns, lingering gaps: most organizations fear AI-driven cyberattacks but lack...

Tenable Forecasts Data Security in the Cloud to Take Centre Stage...

Blockchain-Enhanced Cybersecurity-Safeguarding Digital Identities and Data