Understand attention math
Build a tiny LLM
This becomes your robot’s “micro‑cortex”
Add SigLIP/DINO vision tokens
Fuse them into your GPT core
Now you have a toy VLM
Train on robot demos
Predict poses, grasps, locomotion steps
Now you have a toy VLA
Internal chain‑of‑thought
Multi‑step plans
Tool calls (IK, MPC)
Predict outcomes
Simulate actions
Re-plan dynamically
This is the humanoid agentic brain.
On Oct 5, 2026, at 7:31 PM, A J <aj48...@gmail.com> wrote:
--
You received this message because you are subscribed to a topic in the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this topic, visit https://groups.google.com/d/topic/hbrobotics/eaQrucuM9Rk/unsubscribe.
To unsubscribe from this group and all its topics, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/AF3D167D-9DDE-48AB-8842-C04F4F656D6E%40gmail.com.
You received this message because you are subscribed to a topic in the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this topic, visit https://groups.google.com/d/topic/hbrobotics/eaQrucuM9Rk/unsubscribe.
To unsubscribe from this group and all its topics, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/C1EBDC6E-D678-4536-95CC-16EB4A873501%40gmail.com.
You received this message because you are subscribed to the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this group and stop receiving emails from it, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/CAJZ%2Bfr0TQYRZiDTLyxfMJQHmsw12FLui-jdz6cQa_KguFuSuFA%40mail.gmail.com.
On Oct 6, 2026, at 5:49 PM, andy <aj48...@gmail.com> wrote:Thomas,I did some searching, and using a quantized version of an Open VLA it might take 10 - 15 hours to train.Training on a pair of H100 might take about 1/3 the time. But with a tethered arm with vision,
--
You received this message because you are subscribed to a topic in the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this topic, visit https://groups.google.com/d/topic/hbrobotics/eaQrucuM9Rk/unsubscribe.
To unsubscribe from this group and all its topics, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/4F86BA84-478A-46F2-991A-D39DCA4F221E%40gmail.com.
On Oct 6, 2026, at 6:24 PM, Thomas Messerschmidt <thomas...@gmail.com> wrote:Maybe I missed something, what data are you using to train it.
![]() | |
On Oct 6, 2026, at 9:05 PM, andy <aj48...@gmail.com> wrote:Yes, I think at this point the OpenVLA is more than my system can handle. I will continue to researchthe different ways to train robots. A key point seems to be the quality data needed to feed the pipeline.
--
You received this message because you are subscribed to a topic in the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this topic, visit https://groups.google.com/d/topic/hbrobotics/eaQrucuM9Rk/unsubscribe.
To unsubscribe from this group and all its topics, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/890B3CA1-1FEE-4AED-A43C-8A144FA921E0%40gmail.com.
On Oct 6, 2026, at 11:11 PM, andy <aj48...@gmail.com> wrote:My first impression was that this was a fusion of Vision, Language, and Action (VLA).But after searching some, it is described as a high-level policy generator.I asked Search to explain how VLA works when the bot is asked to pick up the apple.
On Tue, Oct 6, 2026 at 10:06 PM Chris Albertson <alberts...@gmail.com> wrote:On Oct 6, 2026, at 9:05 PM, andy <aj48...@gmail.com> wrote:Yes, I think at this point the OpenVLA is more than my system can handle. I will continue to researchthe different ways to train robots. A key point seems to be the quality data needed to feed the pipeline.That was my conclusion too.What I am looking at now for humanoid walking is a “MLP with history buffer”. MLP is the old multilayer perceptron with an input and output layer and some hidden layers, all fully connected.In the top go the current robot state, joint angles, velocity, camera data, and actions come out the bottom. But the history buffer adds most of the previous N states. Then the model knows past joint angles back to some milliseconds in time. Every time you run it, you move the current state to the past, and the oldest one falls off the end. This can run faster, and like a VLA it predicts the next action given the last N states. This is basically like time series forecasting.Years ago, people used MLP but have moved to transformers now that they have big data centers. But I think for a task as simple as walking, MLP should work.--
You received this message because you are subscribed to a topic in the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this topic, visit https://groups.google.com/d/topic/hbrobotics/eaQrucuM9Rk/unsubscribe.
To unsubscribe from this group and all its topics, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/890B3CA1-1FEE-4AED-A43C-8A144FA921E0%40gmail.com.
--
You received this message because you are subscribed to the Google Groups "HomeBrew Robotics Club" group.
To unsubscribe from this group and stop receiving emails from it, send an email to hbrobotics+...@googlegroups.com.
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/CAJZ%2Bfr2aQG1w%3DkdB%3DgbYCFuCgqMFAsjwcuSZCewG%3DOoLN-fLQg%40mail.gmail.com.
<flowchart_vla_Bot.pdf><bot_pickup_apple.pdf>
To view this discussion visit https://groups.google.com/d/msgid/hbrobotics/E42C5848-5E8F-49E6-8B01-867C834E8543%40gmail.com.