Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • TRADING
  • SUBMIT
Crypto Flexs
  • DIRECTORY
  • CRYPTO
    • ETHEREUM
    • BITCOIN
    • ALTCOIN
  • BLOCKCHAIN
  • EXCHANGE
  • TRADING
  • SUBMIT
Crypto Flexs
Home»ADOPTION NEWS»Generative AI enhances robots’ reasoning and action capabilities with ReMEmbR.
ADOPTION NEWS

Generative AI enhances robots’ reasoning and action capabilities with ReMEmbR.

By Crypto FlexsSeptember 24, 20243 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
Generative AI enhances robots’ reasoning and action capabilities with ReMEmbR.
Share
Facebook Twitter LinkedIn Pinterest Email

Lawrence Jengar
24 Sep 2024 07:06

NVIDIA’s ReMEmbR integrates generative AI, vision language models, and augmented search generation to enhance the reasoning and action capabilities of long-term robots.





According to the NVIDIA Technology Blog, NVIDIA has unveiled ReMEmbR, a groundbreaking project that leverages generative AI to enable robots to reason and act based on expanded observations.

Innovative Vision Language Model

Visual Language Models (VLMs) combine the powerful language understanding of basic large-scale language models (LLMs) with the visual capabilities of visual transformers (ViTs). These models can process unstructured multimodal data, infer it, and return structured outputs by projecting text and images into the same embedding space. Based on extensive pretraining, VLMs can be adapted to a variety of vision-related tasks through new prompts or parameter-efficient fine-tuning.

ReMEmbR: Improving Robot Perception and Autonomy

ReMEmbR integrates LLM, VLM, and augmented generation (RAG) to enable robots to reason and act based on what they observe over long periods of time, from hours to days. The system is designed to address challenges such as large-scale context processing, reasoning about spatial memory, and building prompt-based agents that query for additional data until the user’s question is answered.

The memory construction phase of this project uses VLM and a vector database to build a long-horizon semantic memory. In the query phase, the LLM agent infers on this memory. ReMEmbR is completely open source and runs on the device, making it accessible to a wide range of applications.

Real-world applications and demos

To demonstrate the capabilities of ReMEmbR, NVIDIA developed a real-world example using Nova Carter and NVIDIA Isaac ROS. A robot equipped with ReMEmbR can answer questions and guide individuals within an office environment. The demo highlights the system’s ability to build an occupancy grid map, run a memory builder, and operate ReMEmbR agents.

In the demo, the robot uses a monocular camera and global position information to create a vector database. This database stores text embeddings, timestamps, and pose information, allowing the robot to efficiently query and retrieve information to perform tasks such as guiding a user to a specific location.

Integration with speech recognition

Recognizing the need for intuitive user interaction, NVIDIA has integrated speech recognition into the ReMEmbR system. Using the WhisperTRT project, which optimizes OpenAI’s Whisper model with NVIDIA TensorRT, robots can process voice queries and generate appropriate responses to enhance the user experience.

Future outlook

ReMEmbR’s innovative approach of combining generative AI, VLM, and RAG opens up new possibilities for robotics applications. By giving robots the ability to reason and act based on extended observations, this technology has the potential to revolutionize areas such as autonomous driving, surveillance, and conversational assistance.

For those interested in exploring generative AI in robotics, NVIDIA offers a wide range of resources and documentation through its Developer Program, including tutorials, code samples, and community support to help developers get started with their own generative AI robotics applications.

Image source: Shutterstock


Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Related Posts

Stellar (XLM) Highlights the Superiority of Native Tokenization in Securities

May 6, 2026

Bitcoin is at risk of liquidation of $1.4 billion if BTC rises to $80,000.

April 28, 2026

Polymarket Seeks $400 Million Raise to $15 Billion Valuation: Report

April 20, 2026
Add A Comment

Comments are closed.

Recent Posts

Bitmine Immersion Technologies (BMNR) Announces ETH Holdings Reach 5.28 Million Tokens, And Total Crypto And Total Cash Holdings Of $12.6 Billion

May 18, 2026

How to Bet Safely with Crypto: The Most Trusted Licensed Sportsbook

May 18, 2026

Lock.com Enters Early Access With Isolated Signing And Post-Quantum Architecture

May 18, 2026

1win Crypto Tournaments Go Global With Up To 200K USDT In Rewards

May 18, 2026

Ethereum Triangle Breakdown Adds Pressure to Recovery Prospects

May 18, 2026

AFX Launches Sovereign Layer 1, Providing An Optimized Execution Environment For On-chain Perp DEXes

May 18, 2026

DOGEBALL Tracks 2900% Profits, Breaks Poly Truth Capital, Meme Punch Stagnation, Positions itself as Best Cryptocurrency Presale to Buy Now

May 18, 2026

Ripple (XRP) tests $1.43 support amid mixed market sentiment.

May 17, 2026

With Ethereum price stuck below $2,320, hopes for recovery are starting to fade.

May 16, 2026

Washington DC Summit As Real Estate Tokenization Enters Its Next Phase

May 15, 2026

Could BNB price fall above $750 if a double bottom pattern forms?

May 15, 2026

Crypto Flexs is a Professional Cryptocurrency News Platform. Here we will provide you only interesting content, which you will like very much. We’re dedicated to providing you the best of Cryptocurrency. We hope you enjoy our Cryptocurrency News as much as we enjoy offering them to you.

Contact Us : Partner(@)Cryptoflexs.com

Top Insights

Bitmine Immersion Technologies (BMNR) Announces ETH Holdings Reach 5.28 Million Tokens, And Total Crypto And Total Cash Holdings Of $12.6 Billion

May 18, 2026

How to Bet Safely with Crypto: The Most Trusted Licensed Sportsbook

May 18, 2026

Lock.com Enters Early Access With Isolated Signing And Post-Quantum Architecture

May 18, 2026
Most Popular

Blur price doubles after Season 2 airdrop and Binance listing.

November 25, 2023

Bitbot profits as Ape Terminal cancels ZKasino IDO.

April 21, 2024

AssemblyAI Releases C# .NET SDK and New AI Tutorials

September 8, 2024
  • Home
  • About Us
  • Contact Us
  • Disclaimer
  • Privacy Policy
  • Terms and Conditions
© 2026 Crypto Flexs

Type above and press Enter to search. Press Esc to cancel.