/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Nvidia Research announces Eureka, an AI agent powered by GPT-4 that autonomously writes reward algorithms to teach robots to perform complex skills like a human

Sharon Goldman / VentureBeat :

VentureBeat Sharon Goldman

Discussion

  • @drjimfan @drjimfan on x
    Can GPT-4 teach a robot hand to do pen spinning tricks better than you do? I'm excited to announce Eureka, an open-ended agent that designs reward functions for robot dexterity at super-human level. It's like Voyager in the space of a physics simulator API! Eureka bridges the... …
  • @arankomatsuzaki Aran Komatsuzaki on x
    Eureka: Human-Level Reward Design via Coding Large Language Models Demonstrates for the first time a simulated 5-finger shadow hand capable of performing pen spinning tricks at human speed proj: https://eureka-research.github.io/ repo: https://github.com/... abs: https://arxiv.or…
  • @drjimfan @drjimfan on x
    Eureka achieves human-level reward design by evolving reward functions in-context. There are 3 key components: 1. Simulator environment code as context jumpstarts the initial “seed” reward function. 2. Massively parallel RL on GPUs enables rapid evaluation of lots of reward... [v…
  • @scobleizer Robert Scoble on x
    My high school friend constantly did this with his pen. If your robot can do this it can pretty much do anything with your hands. My friend tried to teach me to do it. I never was able to.
  • @chris_j_paxton Chris Paxton on x
    This is a bit buried but I think it's the most interesting bit here. Using reward reflection, they're taking *some* of the prompt engineering out of works like ProgPrompt, Code-as-policies: robotics papers which generated code to solve specific problems Really cool work!
  • @jasonma2020 Jason Ma on x
    Super excited to share Eureka, our “spin” on how to use LLMs to teach low-level dexterity skills! Eureka is an open-ended reward design agent that can write and evolve superhuman reward functions for a large suite of robots and tasks, including challenging pen spinning tricks!
  • @nvidiaaidev @nvidiaaidev on x
    🎉Just released: Eureka!, a new AI agent that uses LLMs to automatically generate algorithms to train robots to accomplish complex tasks. 👀 The #NVIDIAResearch paper includes the AI algorithms and how to experiment with Eureka using NVIDIA Isaac Gym. 👇 https://blogs.nvidia.com/...…
  • @benbajarin Ben Bajarin on x
    The training of robots begins.
  • @drjimfan @drjimfan on x
    Finally, Eureka relies on reward reflection, which is an automated textual summary of the RL training. This enables Eureka to perform targeted reward mutation, thanks to GPT-4's remarkable ability for in-context code fixing. Here we provide an illustrative example of the various.…
  • @guanzhi_wang Guanzhi Wang on x
    Excited to share Eureka, an open-ended agent capable of writing reward functions for robots at super-human level. Great work led by @JasonMa2020!
  • r/singularity r on reddit
    Eureka!  NVIDIA Research Breakthrough Puts New Spin on Robot Learning