• About Us
  • Disclaimer
  • Contact Us
  • Privacy Policy
Wednesday, September 16, 2026
mGrowTech
No Result
View All Result
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions
No Result
View All Result
mGrowTech
No Result
View All Result
Home Al, Analytics and Automation

Generalist AI Releases GEN-1.5: A Robot Foundation Model That Learns New Tasks From One 3–12 Second Demo

Josh by Josh
August 24, 2026
in Al, Analytics and Automation
0
Generalist AI Releases GEN-1.5: A Robot Foundation Model That Learns New Tasks From One 3–12 Second Demo

[ad_1]

Generalist AI has released GEN-1.5, a robot foundation model that learns a new physical task from a single demonstration. Drop 3–12 seconds of sensorimotor data into its 30-second context window, and the robot performs the task. No gradient updates, no fine-tuning, no task-specific programming. Across 10 diverse manipulation tasks, this one-shot in-context prompting averaged 59% success (±10% std. dev.) straight from the pretrained model. Ten gradient steps on five minutes of data per task raised that to 83% (±9%). Generalist calls the mechanism physical prompting, and says it was never trained for: no architectural changes, no meta-learning loop, no auxiliary objectives. It emerged from over eight months of continuous pretraining on physical interaction data. The tasks are simple and short-horizon, and the company says so plainly. But this is the first model its team knows of where one-shot learning of physical skills has emerged at scale.

Is it deployable?

Not yet — this is a research release. There are no public weights, no API, no pricing page and no self-serve product. Generalist AI runs GEN-1.5 on its own fleet and data engine. Anyone who wants it today goes through a direct partnership.

READ ALSO

Cohere Releases North Small Translate: A 218B MoE Translation Model That Scores 83.6 on WMT26 Across 50 Languages

OpenAI Launches ChatGPT for Financial Services With Built-In Data – Unite.AI

What is GEN-1.5?

GEN-1.5 is a large multimodal model that takes video, sensor, language and proprioceptive inputs, holds 30 seconds of memory, and emits 100 Hz action trajectories. It has been pretraining continuously for over eight months on physical interaction data captured in homes, warehouses and factories.

The main mechanism is physical prompting. A sensorimotor example — sensor streams plus the action trajectory — is inserted into the 30-second context window through a drag-and-drop interface. The remainder of the window holds rolling observations. The model then performs the task immediately, with zero gradient steps and no fine-tuning.

Crucially, none of this was designed in. Generalist states there were no architectural changes to promote in-context learning, no meta-learning loop, and no auxiliary objectives encouraging improvisation. The capability emerged from pretraining scale, the same way one-shot prompting emerged in GPT-3.

The numbers

Across 10 diverse tasks, one-shot in-context prompting averaged 59% success (±10% std. dev.) from the pretrained model, with no training at all. Ten gradient steps on five minutes of data per task — roughly 50 demonstrations — raised that to 83% (±9%). In the extreme case, one gradient step on one minute of data reached 66.5% on a held-out task, with no adaptation-specific hyperparameter sweep.

The compute story is the interesting part. Adapting robot policies has typically taken tens of thousands of gradient steps. Ten steps here move the model weights on held-out tasks by less than 0.15%, which suggests fine-tuning is reconfiguring knowledge the model already has rather than building new representations. Generalist frames it as test-time training in an extremely low-data regime.

Three transfer results worth knowing

  • Compositional generalization: Two independently recorded prompts placed in context get chained into one continuous behaviour. The model produces the bridging motions — repositioning, regrasping, error recovery — that appear in neither demonstration.
  • Zero-shot sim-to-real: A demonstration recorded entirely in simulation works as a prompt for the real robot, despite pretraining containing no simulation data — neither rendered video nor simulated dynamics. For some tasks, demonstrations no longer need to be collected physically.
  • Human-to-robot imitation: In some cases a person demonstrates with their own hands, in view of the robot’s cameras, and the model reproduces it with the robot’s hands.

Generalization also shows up after light fine-tuning. Trained on five minutes of brushing a block into a bowl, the model used a banana as a makeshift brush, and used a dustpan to lift and dump the block instead — a different contact sequence entirely. It also removed a sheet of paper covering the bowl, and worked ambidextrously when demonstrations used one hand.

Interactive explainer

Key Takeaways

  • GEN-1.5 learns new manipulation tasks from a single 3–12 second demonstration dropped into its 30-second context window.
  • One-shot in-context prompting hit 59% across 10 tasks; 10 gradient steps on 5 minutes of data hit 83%.
  • One-shot, sim-to-real and human-to-robot transfer emerged from pretraining — none of it was explicitly trained for.
  • Ten gradient steps change weights by under 0.15%, collapsing per-task adaptation compute by orders of magnitude.
  • No weights, no API, no product: treat this as a research signal about scaling, not a deployable system.

Check out the GEN-1.5 research post and @GeneralistAI announcement thread. Feel free to check out our GitHub Page for Tutorials, Codes and Notebooks.

Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.? Connect with us


Asif Razzaq is the CEO of Marktechpost Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.

[ad_2]

Source_link

Related Posts

Cohere Releases North Small Translate: A 218B MoE Translation Model That Scores 83.6 on WMT26 Across 50 Languages
Al, Analytics and Automation

Cohere Releases North Small Translate: A 218B MoE Translation Model That Scores 83.6 on WMT26 Across 50 Languages

September 11, 2026
OpenAI Launches ChatGPT for Financial Services With Built-In Data – Unite.AI
Al, Analytics and Automation

OpenAI Launches ChatGPT for Financial Services With Built-In Data – Unite.AI

September 11, 2026
DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse
Al, Analytics and Automation

DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

September 10, 2026
Security Video Annotation Guide: GDPR-Compliant Labeling
Al, Analytics and Automation

Security Video Annotation Guide: GDPR-Compliant Labeling

September 10, 2026
Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI
Al, Analytics and Automation

Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI

September 10, 2026
MIT Schwarzman College of Computing launches pilot to help educators teach AI across disciplines | MIT News
Al, Analytics and Automation

MIT Schwarzman College of Computing launches pilot to help educators teach AI across disciplines | MIT News

September 10, 2026
Next Post
Birdfy Nest Duo Review: My Own Private Nature Documentary

Birdfy Nest Duo Review: My Own Private Nature Documentary

POPULAR NEWS

Trump ends trade talks with Canada over a digital services tax

Trump ends trade talks with Canada over a digital services tax

June 28, 2025
15 Trending Songs on TikTok in 2025 (+ How to Use Them)

15 Trending Songs on TikTok in 2025 (+ How to Use Them)

June 18, 2025
Communication Effectiveness Skills For Business Leaders

Communication Effectiveness Skills For Business Leaders

June 10, 2025
Comparing the Top 7 Large Language Models LLMs/Systems for Coding in 2025

Comparing the Top 7 Large Language Models LLMs/Systems for Coding in 2025

November 4, 2025
App Development Cost in Singapore: Pricing Breakdown & Insights

App Development Cost in Singapore: Pricing Breakdown & Insights

June 22, 2025

EDITOR'S PICK

Jay Bavisi, Group President, EC-Council – Interview Series

Jay Bavisi, Group President, EC-Council – Interview Series

May 27, 2025
How I learned to stop guessing and just have a conversation

How I learned to stop guessing and just have a conversation

September 19, 2025
The 157 Best Cyber Week Deals—Save up to 57% Off Gear We Love

The 157 Best Cyber Week Deals—Save up to 57% Off Gear We Love

December 2, 2025
Is There an Actual Difference?

Is There an Actual Difference?

June 23, 2026

About

We bring you the best Premium WordPress Themes that perfect for news, magazine, personal blog, etc. Check our landing page for details.

Follow us

Categories

  • Account Based Marketing
  • Ad Management
  • Al, Analytics and Automation
  • Brand Management
  • Channel Marketing
  • Digital Marketing
  • Direct Marketing
  • Event Management
  • Google Marketing
  • Marketing Attribution and Consulting
  • Marketing Automation
  • Mobile Marketing
  • PR Solutions
  • Social Media Management
  • Technology And Software
  • Uncategorized

Recent Posts

  • Cohere Releases North Small Translate: A 218B MoE Translation Model That Scores 83.6 on WMT26 Across 50 Languages
  • The Changing Role of Digital PR in AI Search Landscape
  • How These XL Phones Compete
  • Corporate Event Registration Software: A Practical Guide
  • About Us
  • Disclaimer
  • Contact Us
  • Privacy Policy
No Result
View All Result
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions