❌

Normal view

Walden Robotics Partners With Toyota on Practical Humanoids

3 August 2026 at 16:05


For a while there, it seemed as though robotics as a whole was stuck in a mad rush towards building humanoid robots mostly because it was very possible (and very lucrative) to do so, even without near-term goals that were necessarily realistic. Some of the magic of those first couple of years of the humanoid explosion has stuck around, but there’s also been an industry-wide sobering leading to pointed questions about practicality and value. In other words, starting a commercial humanoid company now is a much different proposition than it would have been just a few years ago.

On 15 July, Walden Robotics emerged from stealth with US $300 million in funding at a valuation of $1.1 billion. Walden is a spinout of Toyota Research Institute (TRI), and it’s spent the last 10 or so years working on hard problems in robotics with the goal of transitioning from research to real-world applications. That seems like the amount of time and experience that it might reasonably take to develop a practical and value-driven approach to deploying general-purpose humanoid robots, and Walden has chosen an excellent starting point by skipping the legs.

“It’s ironic,” says Walden cofounder and CEO Russ Tedrake. “I thought about legs for 20 years; that’s the class I teach at MIT. There are many reasons to build a robot with legs. But the question is, what’s the addressable market? And what percentage of it is covered by a wheeled base?” It’s this focused, practical thinking that sets Walden somewhat apart from many (if not most) of the other companies in this space. Rather than developing a robot first and searching for a viable commercial use case second, Walden instead identified applications where robots can provide value now, and designed a robot that could safely and efficiently meet those needs.

Walden Robotics

Walden Robotics’ Manufacturing Focus

A smiling man in a blue shirt Russ Tedrake is the CEO and cofounder of Walden Robotics.Walden Robotics

Tedrake is light on the details about what specific applications Walden is targeting at this point (citing confidentiality with current commercial partners). Manufacturing and logistics environments where there are a lot of relatively simple and repetitive tasks that aren’t friendly to conveyor belts and preprogrammed robot arms are a good bet. Even in these environments, however, robots still have to find a useful niche because they’re going up against human workers who are more flexible while also cheaper to employ. So the question is: How do you make an argument to a customer that a robot is actually a better solution than their existing human workers?

“You need to find applications with high utilization—where the robot is used 24 hours a day, 7 days a week,” says Tedrake. “Manufacturing is a global imperative right now, and it makes the economics work.” Economic viability is a necessary condition, but it’s not a sufficient one for Walden, or for their partnership with Toyota. People are a big part of Walden’s plan, too.

One of Walden’s major strengths is the company’s partnership with Toyota, which is not all that surprising given that Walden is a spinout from TRI, which is Toyota’s Silicon Valley–based R&D arm. “Toyota was very proud of the work we had done at TRI, and was ready to go big in this space,” says Tedrake. “Part of the excitement of having Toyota as a partner is that their culture is deeply people-first. When talking to Toyota’s leadership, I was never asked how much money this is going to make, but I was asked how it will improve the quality of life for all people.”

A white and orange humanoid robot manipulates an object in its two-finger grippers. The robot’s chonky design allows it to meet the high-payload requirements of useful manufacturing work.Walden Robotics

In this context, at least in the short term, Walden’s approach to improving the quality of life for people is to take over those aforementioned repetitive manufacturing tasks with robots. Tedrake hopes that this will lead to workplaces where skilled craftspeople are able to do even more with their hard-earned expertise, increasing their efficiency, productivity, and happiness all at the same time—a noble goal, although there’s only so much Walden itself can do to make this happen, and not all customers will share Toyota’s priorities.

Wheeled Humanoid Robots in Factories

Many other humanoid robotics companies are also targeting these logistics and manufacturing spaces with general-purpose robots, and they’re doing so by making robots that are as humanlike as possible. The theory is that a humanoid form factor is necessary when operating in human environments. And there are certainly arguments in favor of a humanoid with legs—stairs exist, for one, and legged robots have a smaller footprint compared with ones that have wheels.

But a large wheeled base offers some significant advantages, as Tedrake points out. You’re incentivized to cram the base full of batteries, since more weight near the floor keeps the robot stable, which also solves the problem of running out of power during the middle of the workday. More importantly, a statically stable robot that moves around on a wheeled base can bypass the safety challenges that are currently keeping legged humanoids physically separated from real humans—most prominently, the fact that legged robots can fall over. “Factories already have autonomous mobile [wheeled] robots,” explains Tedrake. “They already have safety cases built around AMRs. You can piggyback on that with a wheeled base.”

A close-up of a robotic gripper grasping a metallic object in a vise. Simple, rugged grippers make the robot suitable for commercial deployment.Walden Robotics

Walden’s perspective on manipulation is similar. Many humanoid companies are using five-fingered hands that are highly dexterous but also highly complex, which Tedrake believes is not a pragmatic approach in the context of commercial deployments. “There’s a question of what you need to do the tasks, but the real question is just durability,” Tedrake says. “We have been deployed in a Toyota factory, and at the end of the week, the hands take a beating, so we built hands that can take that. I have not seen a more dexterous hand that could have done the work our hand has done.”

Walden’s long-term plan is to build “general-purpose robots.” It’s not always clear what a general-purpose robot is, because (I would argue) nobody is quite sure what “general purpose” means. It’s certainly not referring to robots that can do everything; I think the closest we can get are robots that can be taught to do a useful number of different skills, which is why I prefer the term “multipurpose.” It’s a little pedantic, I know, but I think the distinction is important because it moderates expectations in the near term.

Part of where Walden’s optimism towards general purposeness comes from is TRI’s earlier research on diffusion policy, which helps robots learn new skills more quickly by leveraging previously learned skills as a foundation. “Fundamentally, multitasking is a way to get to a general-purpose robot,” Tedrake says. “I believe there is a single platform that can do a lot of tasks that are of high value for real customers. That will give us the experience we need to give birth to this deployed general-purpose capability.”

Developing Healthcare Robotics with GPU-Native Medical Physics Simulation

28 July 2026 at 20:49
A surgeon using simulation on a computer to place a catheter.Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation....A surgeon using simulation on a computer to place a catheter.

Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation. Every demonstration requires specialized equipment, clinical expertise, and access to patients or laboratory environments. This creates three fundamental challenges for developers. First is the data gap. Training modern robotic policies…

Source

💾

Robot Finger Feels in Color

28 July 2026 at 15:00


Imagine running your fingertip over the surface of a U.S. penny. You would feel the ridges of the raised letters and numbers, Abe Lincoln’s bearded side profile, and, if it’s tails, the fluted columns of the Lincoln Memorial. Getting a robot to sense the same things is an imposing task, often requiring gathering data on pressure and force at many spatial locations at once. But that’s just what a team of scientists in Europe has now managed to do, using an unusual, colorful robotic skin that provides high-resolution sensing in real time.

“To be honest, when they showed us this, we thought it was, and pardon my French, [expletive] cool, because it’s a distinctly different approach,” recalled Rich Walker, director of Shadow Robot, the U.K.’s longest-running robot company, which primarily focuses on robotic hands.

RELATED: “This DIY Bipedal Robot Used Pneumatic “Air-Muscles” Instead of Motors”

The research team, which hails from Queen Mary University of London, the University of Florence, the University of Trieste, and the University of Trento, designed a robotic fingertip with a synthetic skin that reflects different colors of light in response to mechanical deformation. By reading the light reflected off the skin, the fingertip generates maps of topology, strain, and contact pressure. The team has already used the sensor to generate maps of a human fingertip, a penny, and a leaf.

Giacomo Sasso, a postdoctoral research associate in the lab of Federico Carpi at Queen Mary University of London, came up with the idea for the sensor. He had been researching optics when he stumbled upon an interesting paper published in the journal Nature. It described the “mechanochromic material” that would eventually make up the reflector in the skin.

Following the method described in Nature, Sasso exposed a light-sensitive film to a 5-megawatt, 635-nanometer (red) laser for seven minutes. The laser beam creates an interference pattern which causes the film to polymerize in alternating densities, creating layers with different refractive indices.

This structure is called a Bragg reflector. The alternating densities and refractive indices in the polymer cause specific wavelengths of light to be reflected. When the reflector is deformed by contact with an object, its layers are stretched, becoming thinner and reflecting light of a different wavelength.

It took Sasso less than a week to re-create the material in the lab. “From there, we started seeing how we could translate these color patterns into something that was useful for us,” he says. Soon, they realized that the color produced by the material was all they needed to be able to sense the topology of objects.

In the robotic finger, the Bragg reflector is sandwiched between a layer of silicone, which protects it from the outside, and a transparent, fingertip-shaped silicone finger with a camera and LED light embedded inside of it.

The light from the LED shines through the clear polymer of the finger. When the fingertip is deformed by an object, the reflector bounces light back to the camera, with wavelengths depending on the level of deformation—red for least deformation, shifting to green, and then to blue when most deformed.

The team also made adjustments to increase the sensitivity of the skin and help the camera to better read color differences. The silicone of the outer layer of the fingertip is colored black to increase the color contrast, allowing the camera to better translate color into the morphology. The rigidity of the camera inside the finger also enhances the deformation of the reflector, producing greater differences in reflected wavelengths.

After all that optimization, the finger provided 100-micrometer resolution with no computational latency, the researchers determined.

What robot fingers need

Human skin takes in a variety of tactile information in order to successfully move and manipulate objects, including temperature, texture, pressure, and vibration. But engineering a robot to do the same is challenging because of spatial constraints. There often isn’t enough room in a robotic fingertip to incorporate more than one type of sensor. The question then becomes: Which type of sensor should be used?

“And the answer to that is…that’s a really hard question. No one knows yet,” says Walker. Carpi’s team’s robotic finger is exciting because it presents yet another option for roboticists to experiment with, Walker says.

Although Carpi’s team isn’t the first to use soft materials for tactile sensing, its technology is unique because it is able to extract quantitative information about depth and size from the topologic maps it generates. According to Walker, most sensors can only generate topological maps, which reveal the relative sizes of object features.

To Sasso, another key advantage of this robotic finger is that it embeds tactile sensing directly into the material of the finger, rather than using taxels, or pixels that measure force or pressure at specific spatial points.

RELATED: Robot Hand Manipulates Complex Objects by Touch Alone

“The core aspect of the sensor is that we’re essentially [moving toward] having the sensing element at the material level,” he says. “The camera, which is a very highly optimized electronic component, is translating whatever the material is already doing directly into digital signals.”

Michael Wang, co-founder and chief scientist at Daimon Robotics, which, unlike Shadow Robot, primarily uses vision-based sensing, echoed Walker’s sentiment that it’s beneficial to explore new methods of sensing, which may bring unique advantages. But he also explained that soft materials often face challenges with durability, and that the significance of the team’s work would be revealed when the finger is integrated into real robot hands.

“The practical and useful benefits, especially in the context of robot hands, remain to be tested and validated,” he says.

When the materials of soft sensors, like the silicone in Carpi’s team’s fingertip, become eroded or damaged after repeated use, the signals measured by the sensors may not reflect objects’ topography as well.

“Especially if you have the electronics embedded into the material layer itself, that becomes a very challenging engineering problem. And I haven’t seen [many] good soft electronics materials that really can undergo long periods of usage,” Wang says.

However, because the Bragg reflector isn’t in direct contact with objects itself, the outer layer of silicone material acts as a protective barrier, Sasso says. The silicone can also be made more durable using certain chemical coatings, according to Wang.

The team has already been talking to companies that could potentially employ the new sensor. They also hope to improve the sensor so that it can sense objects that don’t lie flat on surfaces. That could open up its use in surgical instruments that require precise contact mapping of tissues and organs, Carpi says.

“There are significant developments that we expect with a clear path toward transition to real world applications,” he says.

How to Make an Invisible Drone

16 July 2026 at 16:09


There are many words that I would never, ever use to describe a drone. Stealthy. Subtle. Whatever the opposite of obnoxious is. Much of this is because of the giant angry bee sound that drones tend to make, but it’s also the way that they look in flight: With uncannily linear movements and an even less canny ability to hover perfectly still, they tend to draw the eye as affronts to nature.

In a paper presented this week at Robotics Science and Systems 2026 in Sydney, roboticists from Northwestern University, Evanston, Ill., demonstrated a drone called Phantom Twist that is essentially invisible to humans, being an order of magnitude more difficult to see in flight than a typical quadrotor. They accomplished this with the aid of computational design, and while the resulting hardware is, I would argue, also an order of magnitude more of an affront to nature than a typical quadrotor represents, it’s pretty amazing how well it works.

Phantom Twist spins so fast, it’s practically invisible.Michael Rubenstein/Northwestern University

The trick here is easy to see, even if the drone isn’t. By spinning in flight at between 15 and 25 hertz, Phantom Twist takes advantage of humans’ decidedly mediocre visual system to turn a solid spinning object into an opaque smear. Human eyes take some amount of time (typically about 100 milliseconds) to integrate what we see before sending the full scene off to our brains for processing. Moving objects can cause problems for this system, because if the movement is fast enough, our eyes are forced to average that motion across the scene, combining it with whatever is in the background and resulting in a transparent blur. This effect is called persistence of vision. For something that spins like Phantom Twist, that motion blur comes from the drone’s rapid rotation, and it works because most of the drone is cleverly designed to be empty space.

Drones that spin in flight are nothing new—we’ve covered a bunch of them in the past, including Picolissimo and any number of samara drones inspired by the spinning flight of maple seeds. What makes Phantom Twist unique, and also very odd, is that the design was computationally optimized for low visibility.

Controlling how drones like this fly

Before we get into that, though, a quick note about how drones like this can even fly controllably, because it’s not at all obvious. With just a single motor and no control surfaces, the only possible control input is through the motor itself, and by pulsing the motor speed up or down at just the right time during each rotation, the drone can translate in any direction. Altitude control comes from changing overall motor thrust, and the drone‘s spinning nature makes it passively stable.

A minimalist drone made of a few thin rods, wires and a miniature circuit board. Carbon fiber rods connect batteries, a controller, some counterweights, and a motor and propeller. The research robot also includes optical tracking tags.Michael Rubenstein/Northwestern University

The bits that you need for this kind of drone include the motor and propeller, a couple of batteries, a controller, some counterweights (which could be replaced with more batteries or payload), 0.8-mm carbon fiber rods to tie it all together, and a connector for the handheld launcher that gets the whole thing up to speed. The actual arrangement of these components is surprisingly flexible, and that’s where the invisibility comes in.

“The design space is high dimensional,” explains Northwestern’s Michael Rubenstein. “It’s very difficult for a human to reason through all the trade-offs between the physical constraints required for stable flight and the visual appearance of the spinning drone, and I don’t think we would have easily arrived at this low-visibility design ourselves.”

The visibility (or not) of Phantom Twist is primarily driven by the extent to which different components line up with each other from the perspective of someone looking at the drone. The more components that line up with each other as the drone flies, the less background you see through the spinning drone, and the more visible the drone becomes. Because you might be looking at the drone from a number of different angles, and also because the drone has to be stable enough for controlled flight, there are a bunch of different things that need to be optimized all at once, which is why computational design is effective here.

Phantom Twist’s final design was generated using an iterative optimizer which had a goal of minimizing a metric called learned perceptual image patch similarity, or LPIPS, while making sure that the design could still physically work. LPIPS is the difference between two images: a background image, and a background image with an overlay of the simulated spinning drone. The smaller that difference is, the more invisible that design is. It’s tricky for a human to consider all of the variables at once, but Rubenstein says that the final design does make intuitive sense, because “the automated pipeline prefers placements where components don’t visually overlap as it spins, or where the components are too close to the center of rotation.”

Two variations of minimalist drones made from a few thin rods, wires and a miniature circuit board. Both are barely visible when in-flight. Two iterations of Phantom Twist drones are shown with their handheld launching mechanisms. The better-optimized version [bottom row] relocates the launcher interface to remove components that are too close to the central axis, making them more visible.Michael Rubenstein/Northwestern University

Out of a starting set of around 20,000 feasible Phantom Twist configurations, the optimized design (the one that you see or don’t see in the pictures and videos) has a LPIPS score of 0.0104. A human-designed Phantom Twist is about twice as visible, with a LPIPS score of around 0.2, and a conventional quadrotor (of the same size) would be over 10 times more visible. And there’s still a bit more optimization that could be done with the electrical wiring as well as increasing the baseline transparency of the components themselves.

Phantom Twist is currently controlled using an optical tracking system, which means that it’s not yet capable of flying outside of a controlled environment. But Rubenstein has built other drones along similar principles in the past, which have successfully flown outside, and he’s optimistic about using those techniques to break Phantom Twist out of the lab. The spinning behavior might even enable some useful sensing capabilities, he says. “An interesting possibility is mounting a camera on the spinning body. As the vehicle rotates, it could capture imagery in every direction, effectively creating a 360-degree view of its surroundings that could be used for onboard navigation and control.”

As for what a drone like Phantom Twist could be used for—assuming that the sound can be mitigated somewhat (and there are potential approaches to making that happen), a stealthy microdrone could do all sorts of things with covert surveillance being the most obvious application. For his part, Rubenstein says that he’s personally excited about the potential for watching wildlife, “where a less-intrusive drone could observe animals while minimizing its impact on their natural behavior.” The elephants in particular would certainly appreciate that.

For a deeper dive into all the particulars of this project, read the paper: Computational Design of a Low-Visibility UAV Using a Human-Aligned Perceptual Metric, by Jingxian Wang, Chen Yu, David Matthews, Emma Alexander, Sam Kriegman, and Michael Rubenstein from Northwestern University, which is being presented this week at RSS 2026 in Sydney.

Develop Lightweight USD Runtimes Faster with AI Agents

15 July 2026 at 21:57
A Gif in a warehouse.OpenUSD is an open, extensible framework that provides a common scene description language for physical AI. It enables teams to bring CAD data, simulation...A Gif in a warehouse.

OpenUSD is an open, extensible framework that provides a common scene description language for physical AI. It enables teams to bring CAD data, simulation assets, and real-world telemetry into a shared, physically accurate view of the world. Until now, building a USD implementation has typically required adapting a large existing codebase— even for teams that need a specific memory footprint…

Source

In a First, a Humanoid Robot Performed Live Surgery Under a Surgeon’s Control

13 July 2026 at 23:10

The robot removed a pig’s gallbladder with standard surgical tools in an ordinary operating room.

Watchers held their breath as the robot made its first incision. Hovering over its patient, an anesthetized pig, with a robotic assistant standing nearby, it navigated to the gallbladder and gently removed it.

The operation marked the debut of humanoid robots in a standard surgical setting. The robot, named Surgie, wasn’t autonomous—it was controlled by an expert surgeon—but the study is a step toward using humanoid robots as collaborators in minimally invasive surgery.

“Remotely operated and autonomous humanoid robots have real potential for amplifying access to critical surgeries to which patients would otherwise not have access,” said study author Michael Yip at UC San Diego.

The study included two successful surgeries. Human surgeons remained on standby for emergencies, but the teleoperated robot completed the task with only minimal intervention.

Feedback from surgeons operating Surgie was positive. They reported less physical strain and frustration, along with better overall performance. But they also pointed to practical problems like intermittent overheating and the need to frequently reposition the robot.

Despite a long road ahead, humanoid robots “have a viable future,” said Yip. “You can imagine these robots being deployed in remote communities where staffing is challenging, or in austere environments like search and rescue scenarios where a massive deployment of field medicine is needed in a short period of time.”

Smooth Operator

Robots have assisted surgeons for years. With a human surgeon at the helm, they excel at delicate procedures requiring precision and dexterity. They’re especially well-equipped for laparoscopic surgery, a minimally invasive technique that uses tiny incisions to reduce pain, speed recovery, and lower the risk of infection.

Despite the promise, surgeons face tradeoffs when they use surgical robots. The robots are highly specialized and often require operating rooms to be redesigned to accommodate them.

A major reason for this is the way they’re built. Intuitive Surgical’s Da Vinci system, for example, uses a robot with multiple arms, each independently controlled from a remote console. Other systems, such as Versius from CMR Surgical, deploy several lightweight independent arms, each attached to a mobile base. The robots have to be carted near the patient.

Surgeons operate all these systems from a console using a magnified, high-definition, 3D view of the surgical field, which is often better than what they’d see with their own eyes. Da Vinci 5 adds sharper visuals and depth perception with two cameras, one for each eye. And because the cameras are held by a robot rather than a human assistant, the image is far more stable.

These platforms are already used in a range of operations. But they have weaknesses. Most require proprietary surgical instruments and methods to make extra space for robot docking and maneuvering during procedures. Staff training adds further complexity and cost, limiting where the systems can be deployed.

Humanoid robots, in contrast, are far more mobile and compact. Their human-like bodies could move through standard operating rooms, use conventional surgical instruments, and potentially be easier to incorporate into existing operating rooms.

The timing may also be right. Recent advances in electric components controlling their motion have made humanoid robots faster and more stable than their awkward, stumbling predecessors. Newer AI systems that predict full-body movement and provide feedback have improved robots’ balance and ability to adjust to real-world complexities. Humanoid robots are already stocking warehouses and winning marathons.

But surgery sets a higher bar.

We still don’t know how close humanoid robots are to meeting the requirements for surgical procedures, wrote the team. That’s what they set to find out.

Hello, Surgie

The new system consists of a surgeon’s control console and the robot itself. The surgeon wears a stereoscopic headset with a magnified 3D view of the surgical field and controls the robot with an input device. The robot translates the surgeon’s commands into movements in real time.

The team chose the commercially available Unitree G1 for the job. Unlike Da Vinci, which was built for surgery, G1 is a more general-purpose humanoid with dexterous wrists and multiple joints. The researchers customized the robot’s hands so that it can rapidly switch between surgical tools. Standing just over four feet tall and weighing roughly 77 pounds, the robot takes up a fraction of the space needed by conventional surgical robots.

Precision is key for laparoscopic surgery. Surgical instruments must pivot around a fixed site at the incision, allowing them to move freely inside the body without stretching or tearing neighboring tissues. After extensively mapping Surgie’s movements, the team identified a safe set-up with enough range of motion for most minimally invasive surgeries.

Surgie passed standard robotics benchmarks evaluating surgical skill for both humans and robots. But the real challenge came next. The team performed two gallbladder removal surgeries in a standard operating room. Both operations followed a typical workflow, with a lead surgeon and an assistant responsible for placing the camera, cleaning lenses, and swapping instruments.

Surgie collaborated with the human assistant to locate, identify, and remove the gallbladder with minimal damage to surrounding tissues, including the liver. During part of one procedure, a second humanoid briefly took over camera handling while the human assistant stepped aside.

Both operations went relatively smoothly. One involved minor bleeding and bile leakage from the gallbladder, but both were easily managed. In interviews, surgeons said controlling humanoid robots felt intuitive, particularly because they had two arms and could use standard surgical tools.

“We were surprised at how well Surgie meshed with our workspace and workflow,” said study author Nikita Thareja.

The system is still in early development. Surgie’s restricted reach required frequent repositioning and recalibration, adding more than three minutes each time. The robot also occasionally needed cooling breaks after overheating. In a real operating room, interruptions like these could increase risk by forcing surgeons to split their attention between the procedure and supervising the robot.

Still, Surgie has a leg up on conventional surgical robots: It can walk. Beyond assisting with an operation, it could potentially fetch surgical tools or help clean operation rooms between procedures.

The team is now refining the system to reduce control lag, particularly during long-distance teleoperation, and exploring ways to safely sterilize—or “scrub in”—a humanoid robot for the operating room.

“Our goal is an operating theater of the future, where humanoid robots and humans work side by side as an integrated team to deliver procedures to those in need, both in traditional hospital settings as well as in non-traditional, field medicine scenarios,” said Yip.

The post In a First, a Humanoid Robot Performed Live Surgery Under a Surgeon’s Control appeared first on SingularityHub.

Video Friday: A World Cup for Robots

10 July 2026 at 16:00


Video Friday is your weekly selection of awesome robotics videos, collected by your friends at IEEE Spectrum robotics. We also post a weekly calendar of upcoming robotics events for the next few months. Please send us your events for inclusion.

RSS 2026: 13–17 July 2026, SYDNEY
Summer School on Multi-Robot Systems: 29 July–4 August 2026, PRAGUE
Actuate 2026: 18–19 August 2026, SAN FRANCISCO
IROS 2026: 27 September–1 October 2026, PITTSBURGH
Humanoids Summit Seoul: 22–23 September 2026, SEOUL

Enjoy today’s videos!

For the first time, two full teams of humanoid robots played an 11-vs-11 soccer match on hardware, bringing one of robotics’ most ambitious long-term visions closer to reality. Never before have two full-sized humanoid robot teams played a soccer game against each other.

[ RoboCup ]

Engineers at MIT and EPFL in Lausanne, Switzerland, have designed a robot that can swim underwater and flap out of the water to continue flying through the air, much like a diving bird. The robot can help scientists study the mechanics that enable these actions in aquatic aviators and may help launch a new class of aerial-aquatic drones and vehicles.

[ MIT ]

We’re excited to announce our breakthrough robotic hands for the NEO platform: hands that match or exceed human-level dexterity, strength, safety, and reliability. Designed from the ground up, these 25-DoF hands combine 25 fully actuated degrees of freedom with a tendon-driven system, rich tactile sensing, and built-in compliance. The result is a hand capable of true in-hand manipulation, precision tool use, and delicate interaction.

[ 1X ]

This match, Tech United played against IRIS at the midsize league at RoboCup 2026 in Incheon, South Korea.

[ Tech United ]

Atlas arrived pitchside at NYNJ Stadium in front of 80,000 people gathered to see Brazil vs. Norway. After performing some of the sport’s most memorable player celebrations, Atlas helped kick off the second half by delivering the match ball!

[ Boston Dynamics ]

Navigating discrete terrain such as stepping stones remains a major challenge for legged robots. Conventional approaches often rely on dense environment reconstruction from cameras or lidar, which can be affected by latency, occlusions, and significant computational overhead. We show that proximity sensors integrated into the bottom of a quadruped’s feet enable safe, terrain-seeking autonomous locomotion.

[ Paper ]

On this holiday, Digit is on grill duty. It turns out precise force control is good for more than payload handling. Happy 4th of July from all of us at Agility.

[ Agility ]

We’ve created GEN-1, our latest milestone in scaling robot learning. We believe it to be the first general-purpose AI model that crosses a new performance threshold: mastery of simple physical tasks. It improves average success rates to 99 percent on tasks where previous models achieve 64 percent, completes tasks roughly 3x faster than state-of-the-art, and requires only one hour of robot data for each of these results. GEN-1 unlocks commercial viability across a broad range of applications—and while it cannot solve all tasks today, it is a significant step toward our mission of creating generalist intelligence for the physical world.

[ Generalist ]

Four years at Figure.

[ Figure ]

Reachy Mini is becoming your real AI companion. The Conversation App makes it able to talk fluently with you, help you with your to-do list, remind you of important tasks, and even chat about music. Long-term memory, voice interaction, always ready to help.

[ Reachy Mini ]

Is this sort of thing now a real job for humanoid robots, then?

[ Unitree ]

Quite a story, but is it a real job?

[ EngineAI ]

If you have a cute animal logo for your research, I will always share it.

[ BIEVR-LIO ]

This is very delicate work, although the real challenge would be picking those nuts out of a jumbled bin full of randomly sized nuts, which is how most of us live our lives.

[ Sanctuary ]

Not for me, thank you, although I’m not saying that most of the other humanoid robots out there are any better looking, fundamentally.

[ UBTECH ]

Robotics professor Dr. Christian Hubicki judges robot soccer skills while knowing very little about soccer himself.

[ ORL ]

In this presentation, Brendan Schulman, vice president of policy at Boston Dynamics, outlines the critical role of government engagement in driving the success of the humanoid robotics industry. He demonstrates how legged robots like the Spot quadruped and Atlas humanoid are moving beyond factory settings to deliver real-world value in infrastructure inspection, industrial manufacturing, and public safety. Schulman highlights the intersection of AI and robotics, showcasing how large behavioral models and reinforcement learning enable robots to navigate slippery floors and autonomously avoid workplace hazards. Ultimately, he calls for a proactive national robotics strategy focused on workforce training, safety standards, and ethical frameworks to support supply chain resilience and global competitiveness.

[ Humanoids Summit ]

IEEE Honors Robotics Pioneer Toshio Fukuda

7 July 2026 at 19:02


Toshio Fukuda has been blazing trails for most of his career. He is considered to be one of the most prolific scholars in robotics, writing more than 2,000 research papers and authoring several books on the field. He’s an influential figure thanks to his pioneering work developing biomedical robotic systems, industrial robots, micro-nano robotics, mechatronics, and AI-driven automation.

Fukuda launched one of the first robotics conferences, the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). It is still popular almost 40 years later.

Toshio Fukuda


Employer

Egypt-Japan University of Science and Technology, in Alexandria

Title

Professor and vice president of research

Member grade

Life Fellow

Alma maters

Waseda University, in Tokyo; University of Tokyo

An IEEE Life Fellow, he is a professor emeritus in the department of micro-nano systems engineering and a visiting professor at Nagoya University, in Japan, where he taught for nearly 25 years. Currently, he is a vice president of research at the Egypt-Japan University of Science and Technology, in Alexandria, Egypt.

Within IEEE, Fukuda has held top volunteer positions including the organization’s highest office: He served as IEEE president in 2020, becoming the first person of Asian descent to hold the role.

He’s a former program director of Japan’s Moonshot program, which by 2050 intends to develop advanced AI robots.

Born in Japan, Fukuda has been recognized by the country for his contributions to science with two of its highest awards: the Medal of Honor with a purple ribbon in 2015 and the Order of the Sacred Treasure in 2022.

IEEE honored him with this year’s Richard M. Emberson Award for “distinguished service advancing the technical objectives of IEEE, especially in the area of robotics.” The IEEE Board-level award is sponsored by the IEEE Technical Activities Board. Fukuda received the award on 24 April at a ceremony in New York City.

As a former IEEE president who has served as a master of ceremonies at several of the organization’s major award events, Fukuda noted that he is more accustomed to bestowing awards than receiving them.

“It’s very interesting to be on the receiving end,” he says.

The journey into robotics research

As a teenager, Fukuda spent his summer breaks teaching himself how to build things including transistor radios and steam engines.

“It was very nice to have a hands-on hobby and make these kinds of things myself,” he says. His experimentation led him to study engineering.

He earned a bachelor’s degree in engineering in 1971 from Waseda University, in Tokyo. He says one of his professors there—Ichiro Kato, regarded as the father of Japanese robotics research—was a good mentor who made a positive impact.

Fukuda’s research interests were robotics and mechatronics, a field that combines robotics, electronics, computer science, and control systems.

He went on to earn a master’s degree and a doctorate in science from the University of Tokyo, in 1971 and 1977. During those years, he also attended Yale, where he conducted research on advanced control theory in 1973.

He reflects fondly on his time at Yale: “It was a very nice environment and a kind of free-thinking atmosphere. It motivated me to study more.”

“IEEE doesn’t care who you are, what you do, what country you are from, or whether you are male or female. IEEE accepts people who have energy and passion.”

While at Yale, Fukuda served as an assistant to his advisor—which led him to consider a career in academia, he says, because he enjoyed the freedom that research work afforded him.

But he realized that such freedom comes with a price. University researchers are expected to raise the money that funds their work. He compares researchers to small-business owners who have to bring in money to keep their enterprise afloat.

That realization led him to select robotics as his field because he intended to develop technologies useful to industry, he says.

After earning his doctorate, he returned to Japan in 1977 to work as a research scientist at the government’s Mechanical Engineering Laboratory, later renamed the National Institute of Advanced Industrial Science and Technology, in Tsukuba.

“There was a lot of research going on at the lab, including practical robotics and theory,” he says.

He left Japan in 1979 to become a visiting research fellow at the University of Stuttgart, in Germany. During his year there, he studied systems, software problems, and related topics.

He returned to Japan and was hired as an associate professor of mechanical engineering at the Tokyo University of Science. He conducted research into practical uses for robots by visiting industrial plants. He decided to develop robots that inspect industrial equipment such as those used in assembly plants, oil refineries, and power stations—places that “can be hostile environments for humans,” he says.

His work drew interest from chemical, oil, and utility companies.

“I got a lot of money from them for this very practical application, which funded my research,” he says, laughing.

Developing popular robotic systems

Fukuda grew tired of making those robots, he says, so he switched to creating ones for scientific applications. He developed many techniques, but he probably is best known for his modular, cellular robotic systems (CEBOTs), which he introduced in 1985.

He has described how CEBOTs work in numerous papers published in the IEEE Xplore Digital Library.

The CEBOT system is composed of a number of autonomous robotic cells that stick together like interlocking Lego plastic bricks, he says.

Each cell is a fundamental modular unit that has a function. When a simple task is given, the system can analyze it and generate the structure of the cellular manipulator. The cells connect to and detach from each other through connection mechanisms and cooperate mutually, creating complex structures and configurations.

“You start developing from the component-wise to the cell-wise to a small functional unit—and then you come up with clusters that make bigger systems. We can make a society of robot beings like that,” he explained in his oral history published on the Engineering and Technology History Wiki. “It’s a distributed robotic system, a self-organized robotic system, and also an evolutionary robotic system.

“It’s also a fault-tolerant robot system because if something is wrong, you just remove those things and make a new one. You keep the system working. That’s a great thing.”

Today CEBOTs are used for a variety of tasks such as delivering medication in hospitals, assisting with planting crops, and transporting products in distribution centers. Check out IEEE Spectrum’s Robots Guide for news from the world of robotics.

In 1989 Fukuda joined Nagoya University as a professor of mechanical engineering and micro-nano systems engineering. During his 24-year career there, he was director of the university’s Center for Micro-Nano Mechatronics. He developed a long list of technologies at the university, including many for medical applications. He also conducted groundbreaking research into intelligent robotic systems and micro- and nano-robotics.

Another technology he is known for is brachiation robots, which he helped develop in 1988. He calls them monkey robots because they’re based on the pendulum-like movement of monkeys swinging from tree to tree. The gravity-based locomotion enables continuous movement.

Brachiation robots now are inspecting high-voltage transmission towers and bridges, searching damaged buildings for survivors, and performing maintenance on pipelines and cables.

Fukuda retired from the university in 2013 and was named professor emeritus.

He didn’t stay retired for long, though. He next held a teaching appointment at Meijo University, in Nagoya, until he left in 2022 to join the Egypt-Japan University.

A prominent volunteer

He joined IEEE in 1980 at the encouragement of one of his research advisors, Professor Fumio Harashima, now an IEEE Life Fellow. After attending conferences and reading the organization’s publications, Fukuda says, he looked forward to becoming more involved.

“I wanted to know how to organize a conference and how to edit a paper for one of its Transactions,” he says. “I wanted to know what was going on from inside the organization, not just the outside.”

In 1988 he was the founding chair and organizer of IROS, in Tokyo. The conference had 330 attendees that year, and was supported by Harashima. Today it is one of the largest and most prestigious conferences on the topic, attracting more than 9,000 people annually. Out of 120,000 conferences, it was the only conference in the Nature Index database for this year, Fukuda says.

In 1996 he and other members launched IEEE Transactions on Mechatronics.

He was the founding president of the IEEE Nanotechnology Council, which was established in 2002. He is considered a pioneer in nanotechnology research, particularly regarding how it relates to robotics.

Over the years, he has held numerous volunteer positions on IEEE editorial boards and committees.

He was the 1998–1999 president of the IEEE Robotics and Automation Society, becoming the first non-U.S. member to hold the title.

He was director of IEEE Division X (2001–2002 and 2017–2018), which covers intelligent systems, biological engineering, robotics, control systems, and photonic technologies. He served as the 2013–2014 director of IEEE Region 10 (Asia-Pacific).

As the 2020 IEEE president, Fukuda saw the organization through the early part of the COVID-19 pandemic. Because of travel restrictions, he realized IEEE should change how it offered its in-person services, specifically educational programs. He encouraged IEEE Educational Activities to develop an online learning platform. The IEEE Learning Network started with just three courses and now offers nearly 2,000 courses, webinars, and learning materials.

An award-winning member

The Emberson Award joins a slew of other recognitions Fukuda has received from IEEE. They include several from the IEEE Robotics and Automation Society: a 2004 Pioneer Award, a 2009 Saridis Leadership Award, and the 2011 Harashima Award for Innovative Technologies. He is also a recipient of the Board-level 2010 IEEE Robotics and Automation Technical Field Award.

He says he feels strongly that IEEE should be a diverse organization that is welcoming to all. As IEEE president, he led efforts to devise a diversity, equity, and inclusion program. Several policies, procedures, and bylaws were revised to give members a safe, inclusive place for discourse.

“It’s important for IEEE to make everyone feel comfortable,” he says. “DEI programs are important. All people should be equal. IEEE doesn’t care who you are, what you do, what country you are from, or whether you are male or female. IEEE accepts people who have energy and passion.

“It accepted me, from the Far East. That’s why I like it.”

You can learn more about Fukuda and his career from the oral history conducted by the IEEE History Center.

Develop Humanoid Robot Policies End-to-End with NVIDIA Isaac GR00T

7 July 2026 at 17:05
As more teams move from humanoid robot bring-up to task-specific skill development, the need for repeatable development workflows is growing. Building humanoids...

As more teams move from humanoid robot bring-up to task-specific skill development, the need for repeatable development workflows is growing. Building humanoids remains complex, and today’s development pipelines are still highly fragmented. As a result, developers spend significant time configuring robotics infrastructure before they can focus on building robot capabilities.

Source

Japan Pioneered Humanoid Robots—Can It Now Catch China?

4 July 2026 at 11:00


“In the future, the relationship between humans and robots will deepen, and the distinction between them will probably disappear.” This prediction, from one of the attendees at the recent Humanoids Summit in Tokyo, might have been unremarkable had it not come directly from an android that was first introduced to the world 20 years ago.

Geminoid HI-6 is the sixth-generation of a robot originally designed in 2006. The mechanical twin of Osaka University professor Hiroshi Ishiguro, Geminoid HI-6 is now equipped with a large language model trained on Ishiguro’s own writings and interviews. It has advanced conversational skills and can even have a chat with its creator, an eerie spectacle. But at the Humanoids Summit, Geminoid was one of the few humanoid robots from Japan, the country that pioneered the form factor.

While the event in Tokyo had only about 40 robots on display, Chinese systems outnumbered Japanese by roughly three to one. Some Japanese robotics firms were even using Chinese robots in their own technology demonstrations, something that would have been unthinkable in the recent past—one Japanese engineer described the situation as “sad.” The conference was a stark reminder of how Japan has ceded its early lead in humanoid robot development to overseas competitors, and the challenge it now faces to secure a place in an ecosystem increasingly dominated by general-purpose robots powered by AI.

Twenty-five years ago, Japan was turning out groundbreaking humanoids that were showstopping in their abilities, but they were not commercialized as practical machines in any meaningful way. Heavily influenced by science fiction and lacking practical applications, they were mostly expensive technology demonstrations that were eventually mothballed. What Japan retains, however, is robotics design and know-how, which it must leverage to be a key player in the rapidly evolving humanoid ecosystem.

Learning to Walk—Then Standing Still

To anyone who has seen recent videos of Chinese humanoids doing kung-fu and synchronized acrobatics, as well as half-marathon races, China’s remarkable progress in the field is nothing new. At the Humanoids Summit, Toyota showed a video of its latest basketball-playing robot, and Honda exhibited its latest robot hand, but the full-scale humanoids on the floor were mostly Chinese–the kid-size K1 machines from Booster Robotics of Beijing were dancing to Michael Jackson tunes. The full-scale G1 humanoid from Unitree Robotics of Hangzhou was also doing demos.

“You cannot sell these bipedal systems in Japan for safety and compliance reasons,” says Shuichi Nagao, a frequent visitor to China as CTO of Omakase Robotics, a division of Zeals, a Japanese humanoid robot developer. Omakase was exhibiting a G1 modified with an external PC controller, a dextrous hand, a suction-cup manipulator and a sensor “hat” with an extra speaker, mic, and camera.

“In China, the government is pushing humanoid development. They didn’t have an industry 20 years ago. The people pushing it are young, in their 20s and 30s. It’s a really different mentality out there,” says Nagao. “Big players in Japan are still looking for use cases for humanoids. In China, they’re already doing mass production and reducing the cost, so other countries can’t compete with them anymore.”

Another Japanese company showing off G1 bots was summit sponsor GMO AI & Robotics, a subsidiary of Japanese internet company GMO. It’s using the robots in partnership with Japan Airlines to load and unload cargo containers at Tokyo’s Haneda airport. The cargo project is a trial—like many other humanoid experiments—but the fact that Chinese machines have penetrated so far into Japan’s ecosystem upends a long history.

In 1973, scientists at Waseda University in Tokyo built WABOT-1, considered the first full-scale humanoid robot, which was capable of slow bipedal locomotion, grasping objects, and simple communication. It inspired Honda’s groundbreaking Asimo humanoid, but Asimo was never commercialized. It was eventually retired in 2022, the year ChatGPT was released. Two years later, Unitree’s G1 went on sale for US $16,000.

A 65 centimeter tall bipedal robot, its design features the head of an anime-style girl with legs directly underneath. China’s High Torque Technology Co. showed off its Mini Pi biped, customized with an anime-inspired head, at Humanoids Summit in Tokyo. The regular version is priced at $3,500. Tim Hornyak

Supply and Demand

Japan’s development of humanoids happened before practical applications or widespread demand were in place, but bad timing is only part of the story—Japan also has a history of developing technologies that might appeal to domestic consumers but not necessarily those overseas. For example, decades after its highly engineered multifunction toilets first appeared, they have only recently found a following abroad.

Japan’s humanoid prowess was partly built on the back of its legendary industrial automation, yet even that stronghold has eroded. Ani Kelkar, a partner from McKinsey & Company in Boston who produces analytical reports about the robotics industry, told the summit audience that while Japan occupied the top spot in the world in manufacturing robot density (the number of multipurpose industrial robots in operation per 10,000 employees) from at least 1994 to 2009, it then slipped to second in 2014, third in 2019, and fifth in 2024. In that year, South Korea was at the top of the leaderboard with a robot density of 1,220 compared to Japan’s 446.

The International Federation of Robotics estimates China now has the most operational industrial robots in the world, with around 2 million total units, approximately 4.5 times more than Japan. “The annual installation numbers are impressive too: 54 percent of all robots installed worldwide in 2024 were deployed in China,” the IFR said in a release in April 2026.

“I think the loss of Japanese leadership is more to do with the rise of China as a manufacturing powerhouse including for sectors that Japan had high export levels,” Kelkar said in an email interview. “The recovery has not yet happened as Japan “missed” the rapid acceleration in AI for robotics and is now playing catch-up.”

How Japan Can Adapt

Kelkar believes Japan has a $100 billion opportunity in general-purpose robotics, which are machines that can perform a wide variety of tasks, and it cannot rely on the slower-growing industrial robot market, which is centered on factory machines that do one simple and predictable task like welding car parts. He points to a McKinsey white paper suggesting that while Japan has much of the hardware and technology experience needed to support general-purpose robot development, it must change its strategy to capture a larger share in AI, software, data collection, and robotics platforms.

Tetsuya Ogata is a professor of engineering and director of the Institute for AI and Robotics at Waseda University, the birthplace of humanoids in Japan. He briefed the summit on how a nonprofit he chairs, the AI Robot Association (AIRoA), is working with Toyota and other members to develop foundational technologies for collaborative use.

For instance, AIRoA has collected some 80,000 hours of data on remote operation of mobile manipulators, which Ogata believes is the largest dataset of its kind. Using the data, it built and verified vision-language-action (VLA) models, and it has also started data collection for dual-arm mobile manipulation. In an interview, Ogata acknowledged Japan’s struggle to find its place in the changing landscape.

“The world of AI is inherently a game of scale,” says Ogata. “Therefore, Japan’s absolute prerequisite is to secure a competitive baseline of scale—in data, computing resources, and talent. Beyond that, what I consider most critical is a mind-set shift: Rather than trying to hoard scale within a single nation or company, we must grow stronger by collaborating with a diverse ecosystem of domestic and international players.”

Specifically, this means creating a “collaborative domain” to address data—the single biggest bottleneck—through industry-wide cooperation rather than data siloing. By collectively nurturing a precompetitive, shared data infrastructure and foundation model, individual companies can then compete on top of it with their own applications. “By offering this open ‘data ecosystem’ to the world, we can engage global players and establish a ‘third pole’ alongside the U.S. and China,” says Ogata. “I believe this is how Japan can reclaim its global presence.”

In 1999, Japan introduced the world’s first mobile internet services platform. But being first didn’t turn Japan into a smartphone manufacturing or design center—it’s now merely a supplier of parts to other countries that are leading the smartphone industry. If Japan can avoid a repeat of that experience and successfully deregulate, diversity, and commercialize its original humanoid dreams, it stands a better chance of influencing the direction of the industry and reaping billions in value. As automobiles and electronics were pillars of Japan’s industrial strategy in the last century, Japan could make humanoid robots one of its key value generators in the 21st century, an approach that would not only deliver economic benefits but give Japan greater clout in how the industry will evolve. Just like Japanese cars, electronics, and even toilets, Japanese humanoids could stand for craftsmanship and reliability. It’s a legacy that Japan can’t afford to give up.

This article appears in the September 2026 print issue as “Japan Seeks a Humanoid Robot Comeback.”

Companies Could Soon Staff ‘Stubbornly Local’ Jobs With Workers 4,000 Miles Away

25 June 2026 at 16:02

Companies once moved whole factories overseas to reduce labor costs. Now, workers a world away can operate local excavators, forklifts, and even humanoid robots with an internet connection.

Packaging potassium sulfate, a fertilizer vital to the planet’s food supply, is visually striking—not because of what you see, but because you don’t see much at all. In China’s Xinjiang region, home to the world’s largest deposit of the mineral, piling it up in warehouses creates dust clouds so severe that workers are forced to drive heavy machinery by feel.

Some companies are now turning to a technology that not only offers a way to see through the dust but also keeps workers from entering the warehouse at all. The system, developed by BuilderX Robotics, a Chinese tech company, uses cameras that are like night-vision for dusty areas. More significantly, operators drive excavators, loaders, and other machines from a remote office filled with rows of videogame-like stations. All they need is a 5G or satellite connection.

The ability to control physical machines from a distance is called teleoperation, and it could become a significant force of change in the global economy.

In Japan, the shelves of over 300 convenience stores are being restocked by robots monitored and sometimes controlled by workers in the Philippines. Düsseldorf airport was slated to begin testing shuttles driven by remote workers in May. A startup in Atlanta is offering robot security guards operated by remote staff, and last summer, a surgeon in France performed a teleoperated procedure on a patient in India.

While offshoring teleoperated jobs to overseas workers hasn’t yet become routine, Mark Graham, professor of internet geography at the University of Oxford, suggests the technology is worth our attention because it might enable companies to expand on their well-established habit of outsourcing jobs to places where labor is cheaper.

The use of remote labor isn’t new, Graham told SingularityHub. But teleoperation extends the logic of outsourcing to tasks that were previously thought to be “stubbornly local.”

“The novelty is less about the existence of remote labor and more about the kinds of work that can now be pulled into a planetary labor market,” he said. “Once that happens you can expect the usual pressures around labor arbitrage, control, and fragmentation to follow.”

It’s not clear we’re ready for the consequences.


BuilderX Robotics is a global leader in teleoperation for heavy machinery and a good expression of the changes ahead. Shaolong Sui, a graduate of Stanford University with a degree in mechanical engineering, founded the company in 2018 as a response to labor shortages in the construction industry in Asia.

“A shortage of trained operators isn’t a problem only in developed countries,” he told me. “Young people here in China don’t want to do this work. It’s dusty and dangerous.”

Rather than focusing on full robotic autonomy, which many construction companies have pursued over the past decade, Sui identified teleoperation as a more realistic way to move operators from harsh environments to safer conditions. Making use of the proliferation of low-cost sensors and 5G at the time, Sui completed a prototype in 2019. Today, his company offers teleoperation for 14 different industrial machines, including excavators, loaders, and bull dozers.

In our conversation, it was clear he hopes to improve working conditions for manual laborers. I lost track of the number of times he mentioned removing operators from dangerous worksites. “These workers deserve a better life,” he said.

BuilderX’s workstations do seem to have transformed some of the punishing work of an industrial site into a more white-collar experience, complete with tea and coffee break rooms and toilets down the hall. Sui said his solution allows construction firms to hire senior citizens or people with disabilities who, thanks to the videogame-like interface, can now operate heavy machinery. In another video, a Japanese woman who pilots an excavator proudly shows off her complex nail art, something she claims she couldn’t maintain when she worked in the field.

“Not only is this a much safer workplace, but the lifestyle benefits are that you can sit in an air-conditioned space, enjoy your tea, and when you go home, you’re still clean,” Sui said.

There’s no doubt the approach is safer for frontline workers like those in Xinjiang. Evidence suggests that high levels of potassium dust exposure can cause chronic bronchitis. While pulling someone from dangerous work is a good thing and that should be taken seriously, Graham told me, it doesn’t necessarily mean they’re free from exploitation.

“A worker can be removed from the physical site and still be subjected to intense surveillance, deskilling, isolation, fragmented contracts, algorithmic management, and downward pressure on wages. In other words, the risk can move rather than disappear,” he said.

Sui and Graham both agree there are plenty of forces that might slow the pace of outsourcing. Currently, none of BuilderX’s customers offshore work to overseas operators. But that doesn’t appear to be a technology constraint, as recently demonstrated by an operator in Poland controlling an excavator over 4,000 miles away in Beijing. On the technical side, latency—the delay between operator and machine—and reliability will shape the rate at which firms can choose to offshore workers. But it’s more likely to be limited by regulatory constraints in the form of licensing, insurance, and safety requirements.

That said, Graham believes the biggest force driving work overseas will be the same one that’s pushed clerical and service work offshore; the relentless pursuit to increase profit and reduce cost.

“If firms can hire people in lower-wage labor markets to operate expensive equipment thousands of miles away, many of them will try,” he said.


Most debates about AI and robotics focus on job loss due to automation. There is relatively little discussion about the risk of offshoring teleoperated work as the technology comes online. This is partly due to the hype surrounding physical AI, a Silicon Valley buzzword describing a world where fully autonomous robots cut humans out of the loop. But Graham says that when machines arrive people tend to incorrectly assume humans disappear.

“In many cases, what gets described as automation is really a reorganization of labor. Work gets broken apart, moved around, and hidden from view,” he says.

As is the case with AI,  the robotics industry’s push toward full automation is still plenty reliant on a hidden system of faraway workers. Teleoperation provides training data for robots and is needed to help them deal with unexpected events. Consumer robotics startup 1X is selling a $20,000 humanoid that will sometimes need to be  controlled by remote staff. It’s not clear how often future robots cleaning dishes in San Francisco kitchens will be steered by gig workers in Mumbai.

Robotaxi company Waymo already relies on human agents to assist, though not literally drive, vehicles stuck in difficult scenarios. The firm recently disclosed for the first time that some of these agents are based in the Philippines. This information, surfaced during US congressional testimony, immediately raised questions of oversight for safety-critical work: For instance, should a worker in Manila be required to get a California driver’s license?

Amid an already combustible US political environment, teleoperation could raise the heat even higher. Fueled by fears of Americans losing jobs to people overseas, Wyndham Hotels and Resorts, the parent company of La Quinta, was last year forced to respond to anger over a viral video depicting workers allegedly in India remotely handling check-in at one of their Miami hotels. As Graham points out, people tend to care more about outsourcing when it’s no longer hidden in a back office.

But outrage alone, he says, rarely defeats a business model that saves money. Due to network effects surrounding training, infrastructure, and other business process optimization, outsourced labor also tends to cluster in specific areas. This may already be happening in the case of Waymo, which could soon see the rise of something like a “driving district” in Manila. In the future, other types of teleoperated work could follow suit, giving companies a ready-made destination to shop for low-cost labor.

For Graham, it’s urgent that we begin requiring certification from independent bodies, which can better scrutinize a company’s production networks. At Oxford he directs Fairwork, a project aiming to improve labor practices in digital supply chains.


I asked Sui how he thinks his customers may reorganize their operations around this new ability to remotely control their machinery.

“We’re working with traditional industries, and so it’s not just about adopting a new technology. There are significant management changes they will have to navigate. You could call this transformation friction because they will need time to digest this new capability step by step,” Sui said.

Despite the fact they could use the technology to outsource work across national borders, none of his customers are doing so just yet. Sui used open pit mines as an example. In this case, where fully developed towns with schools and hospitals have built up over decades, his customers still cluster their workforce next to the sites where they operate. Instead of driving into the mine, operators work from an office and go home clean at the end of a shift.

BuilderX has deployed its technology at more than 100 sites in China, Japan, and parts of Europe. It’s now expanding into new markets including South America and the Middle East. When asked whether he thinks his technology will be used for transnational outsourcing, there’s no hesitation. “Oh yes, I think this is coming in the very near future.”

The post Companies Could Soon Staff ‘Stubbornly Local’ Jobs With Workers 4,000 Miles Away appeared first on SingularityHub.

Inside NVIDIA Halos for Robotics: A Full-Stack Functional Safety System for Physical AI

Physical AI—robots working autonomously alongside people in factories, warehouses, hospitals, and homes—is arriving faster than most expected. Traditional...

Physical AI—robots working autonomously alongside people in factories, warehouses, hospitals, and homes—is arriving faster than most expected. Traditional safety which was built for structured environments can not work anymore as the spaces become more unstructured and robots move out of cages. AI-driven safety is the key. Marking a major milestone in the arrival of physical AI…

Source

What Amazon’s Astro Taught Me About Giving Robots a Soul

19 June 2026 at 10:00


In 2018, Amazon brought me in as the lead UX Sound Designer for Astro, its first consumer home robot. Astro used cameras and other sensors to map and navigate your home and workplace, and could proactively patrol, check up on loved ones, and transport small items using its built-in cargo bin. While there was a well-defined feature set and form factor, initially there was no character direction. In fact, even before Astro had a name, there were two main questions—was it simply Alexa on wheels, or was it a robot with its own character?

The Astro team was divided. One option was to focus on Alexa, and treat the mobile robot simply as an added utility. Along with the majority of the UX team, I argued for Astro to not focus on Alexa. Our belief was that a thing that moves through your home and turns toward you with intent can never be just an appliance. People would ascribe character to it whether we wanted them to or not, and so the only question was whether we shaped that character or let it happen by accident.

Ultimately, Astro became Astro rather than Alexa, and user testing backed up our decision. People didn’t see the robot as Alexa. They saw it as its own character, and that’s what they wanted it to be. Alexa on the device felt somewhat strange and creepy, but building Astro its own voice was too slow and expensive in 2018. So, we settled on Alexa as a supporting character that handled any actual talking, while Astro was the main character, communicating as much as it could without words, through sound, motion, and facial expressions.

I had been brought on to the Astro team to define the robot’s sound design language and voice. But there was no one to flesh out the robot’s actual character. You cannot make a single real decision about a character without defining it first. Every choice about how Astro moved, sounded, paused, or reacted was a character choice, and those choices required all disciplines working together. As sound lead, I was weaving together sound, motion, and character, and how they played together inside each story moment. The animators, who programmed Astro’s motion and facial expressions, were extraordinary at what they did, but the emotional arc they were animating came from the sound (and therefore character) work first. So I stepped into that role, which is where my real work started. What I learned about building character for robots applies to nearly everything being built in embodied AI right now.

Character Is a Design System

Developing a character for Astro meant answering questions that had never been asked about a product at Amazon: What is the emotional range of this robot’s baseline state? How does this robot communicate uncertainty without eroding trust? Where is the line between being expressive and annoying? What are the vulnerabilities of this device’s character?

These are design questions. They have real answers, and every team working on the product has to build from them. For example, Astro’s emotional range was designed to be relatively small at first. We never wanted Astro to get too sad or too angry. It could play sad, but would snap out of it quickly and end the reaction on a high note to keep things positive.

Character leaks out of every seam and can create a disjointed experience if not defined correctly. Even if it’s just animation timing that’s slightly off, or a response that’s technically correct but contextually tone-deaf, users feel every one of these inconsistencies, even if they can’t name them. Watch what happens at the beginning and end of this Sing sequence:

Astro goes from nothing, into the emotional moment, and then lands back on nothing. No buildup, no cooldown, no sense that the feeling came from somewhere or had anywhere to go. I pushed hard for better character stitching, the transitions in and out of expressive moments that make a performance feel continuous rather than assembled, but it never got implemented. The moment itself works. But without the stitching, it reads as a clip playing on a robot rather than coming from within the robot character itself.

Story and Sound at the Beginning

We had decided that Astro would have no spoken dialogue, but it had something that functioned the same way: a vocabulary of sounds, tones, and rhythms that acted as its voice. This vocabulary became the leading output of the character’s personality. The robot’s motion and facial expressions were built around it.

Astro’s wake-up sequence is a great example. Waking wasn’t just a boot animation on the screen; it was an entire performance. Slow and humble at first, the robot oriented itself quietly, then stretched its screen, checked its wheels, and finally, with an upward gesture toward its telescoping mast, it popped it up slightly, and did a little dance of joy. Sound, motion, and eyes hit every beat together in full choreography.

The character’s output in that sequence was first written as a story. Astro is waking up in its new home for the first time. Its main aspiration is to be part of a family, so this is the moment it has been waiting for, this is its purpose. Being the responsible character that it is, it wants to make sure everything is good to go before it introduces itself and starts learning its new home.

This narrative came first because it drove every other decision that we made. After the story was written, sound gave that story a metaphorical voice: the excited tones, the pacing as it checked its wheels, and the bright melodic phrase as Astro looked up at its new family for the first time and introduced itself. Once the sound was laid down, the animation team did their thing with motion and facial expressions, taking cues from the emotional arc the sound had established. Motion didn’t lead—it followed the feeling of the story and the sounds, the same way an animator follows a recorded vocal take.

That wake-up sequence became one of the most-discussed moments in early user testing. People described it as “alive.” What they were responding to wasn’t any single element. It was all three channels (sound, motion, and facial expressions) expressing the same defined character in harmony.

Context Is Where Character Becomes Real

The most compelling characters are defined not by a fixed disposition but by how they respond to their environments and the people in them. They’re still recognizably themselves even as they adapt. This is what I call contextual character. A robot living in a home doesn’t occupy a single emotional state. It moves through rooms with different energy, encounters people in different moods, operates at different times of day, and responds to an endless range of social situations it was never explicitly designed for.

We got close to a contextual character output with Astro’s sound. When a specific piece of environmental context was fed in, the system adapted beautifully, and Astro felt completely alive. But every state like this was still a prediction we made by hand—a situation we had to imagine in advance and design a response for. A random home throws more situations at a robot than anyone can possibly predict, so there was always a longer tail of moments the system was never prepared for.

The difference between a product people describe as “smart” and one they describe as “aware” often comes down to this. Smartness is capability. Awareness is context. Presence is character. And character is always in reaction to the people around it, to its environment, to its own evolving state. That’s what makes it feel like something is emotionally present with you.

This is where AI changes the game for character design in ways that go well beyond what was possible with Astro. AI-driven adaptation doesn’t require the contextual predictions that we relied on. It learns the specific rhythms, preferences, and emotional context of the people it lives and works with. The character doesn’t just respond to context. It grows into it.

What Industry Is Missing

The character and soul of the impending wave of embodied AI products appears to almost always be an afterthought. And character defined late is character defined by default. It becomes the sum of a thousand small decisions made by different people thinking about anything but character. People project character onto devices whether you plan for it or not, especially if those devices move—a robot that moves is already a character. If nobody has designed this character, the result will be products that feel like nothing, or worse, feel confusing and not trustworthy. Technically impressive, but lifeless.

We did not get this fully right with Astro. So many things were moving in parallel that character was rarely treated as a utility, and it made sense why. When you are building a first-of-its-kind product, the things that are the loudest are the ones that break, the deadlines, the costs, the features a customer can point to on a box. Character is quieter than all of that. It’s easy to assume it can come later. On a team as large as the Amazon Astro team, it’s lucky to get any idea onto the road map when it is competing with a hundred others that all feel more urgent in the moment. None of this came from people not caring. It came from character being the kind of thing that is hard to prioritize until you see what its absence costs you.

My Asks to Product Leaders

If you are building a product that will share physical or conversational space with people, three things are worth considering:

Define character before you define interactions. You need a defensible character with enough emotional logic to answer hard questions consistently. Find answers to character questions early, and have every discipline build from the same foundation.

Build story and sound into the character pipeline, not the production pipeline. Story and sound developed alongside character definition has the chance to inform motion, expression, and interaction logic. This requires a different kind of collaboration, and a different kind of hire.

Design for adaptation, not just consistency. A consistent character is necessary, but the products that will matter most in people’s lives are the ones that deepen through use. The infrastructure to support that is more and more accessible, but the design thinking to take advantage of it is still rare.

An expanded version of this story is available on Medium.

The Secret to Marathon-Winning Humanoid Robots

17 June 2026 at 12:19


On 19 April 2026, the Honor Lightning humanoid robot ran a half-marathon in 50 minutes and 26 seconds, beating the human world record by 7 minutes and the best robot time from 2025 by almost 2 hours.

How did Honor do it? Is there some magical technology or technique that unlocked this performance? How did the company beat the significantly better-known Unitree (which reportedly had to supply its robot with an ice backpack to try and complete the race without overheating)? My doctoral thesis involved building and controlling hopping and running robots, and since then I’ve tried to design and build efficient commercial legged robots, giving me a decent idea of the constraints involved. In this article, we take a look at the fundamental underlying constraints to try and answer these questions.

The Physics of Running

Running consists of alternating phases of a leg pushing against the ground (“stance phase”) and the body flying through the air (“aerial phase”). In the aerial phase, the body falls due to gravity, losing vertical momentum. The leg in stance phase pushes against the ground to redirect the vertical momentum upward, while the other leg swings forward to reposition for the next foothold.

Electric motors use energy to produce torque—the higher the torque, the more energy is lost as heat. Adding a gear train after the motor amplifies its torque and reduces its speed. A large reduction helps with torque production, but since the rotor of the motor itself has to spin faster, it becomes very sluggish at accelerating its output. This is obviously bad for the swing phase described above. These competing effects mean that for a particular motor, there is usually a sweet spot for the gear ratio:

A graph showing the relationship between gearing and motor efficiency, with an optimal gearing ratio in the relationship between stance and swing. The power consumed by a robot leg is minimized at an optimal gear ratio (30:1 in this example).Avik De/Datawrapper

How Honor Did It

While the Lightning’s motor specifications are not published, the hip and knee motors roughly have a 110-to-150-millimeter outer diameter. For an approximate set of motor parameters, I looked to the ILM115x25 motor due to its relevant size and detailed specifications.

We can use a simple physics model to estimate the power consumption for running at 7 meters per second (the Lightning’s average half-marathon speed) as gear ratio varies:

A graph showing that optimal gearing for a robot\u2019s motor dissipates the amount of heat that the motor generates.The light blue curve shows how to pick the optimal gearing (45:1). The dark blue curve shows how much heat will be produced in the knee motor, ~150W for the optimal gearing.Avik De/Datawrapper

We see that the drivetrain is not magical: with a gear ratio chosen for this task (we’ll return to this below), the approximate robot power consumption would be a very reasonable 400 watts.

However, the dissipated knee power ( typically the main thermal limiting factor) is approximately 150 W. This is almost an unavoidable consequence—running at human speeds with a humanoid-size robot will inevitably generate this amount of heat! Over a prolonged period, keeping the motor from overheating would be a challenge, but the Lightning has a trick up its sleeve:

According to Honor, the liquid-cooling pipes penetrate deep into the motors like capillaries. The high-power liquid pump has a heat-exchange flow rate of more than 4 liters per minute. Each of the four drive motors in the lower limbs is equipped with an independent liquid-cooling circuit.

Liquid cooling is not new, but it’s definitely not a commodity. It has shown up in research periodically, and on the commercial side Apptronik tried it for a few of its prototypes but (to my knowledge) does not use it on its main Apollo platform. Basic air-convection-based cooling would not continuously be able to extract 150 W out of the knee motor, and so the cooling technology is a key enabler of this type of performance.

Why Others Couldn’t Compete

Why did Honor’s competitors, including more established and widely shipped humanoids such as from Unitree or Agibot, not compete as well?

We can use the same model to generate an equivalent energetics plot for walking at 1.5 m/s, a much more modest but potentially more common activity for a commercial humanoid robot:

A graph showing that robots with gear ratios optimized for running or walking are inefficient when walking or running respectively. The solid and dashed light blue lines show a running-optimized design, while green lines show a walking-optimized design. The optimal ratio for walking is much lower (30:1 vs. 45:1). However, the power dissipated in the knee motor while running [dark blue] is much higher at 30:1 vs. 45:1—the price to pay for running with a walking-optimized design.Avik De/Datawrapper

The plot adds a new green curve for the walking power, and the optimal gearing is significantly different!

Let’s say you design your robot to excel at the normal walking task and choose the green design with 30:1 gearing. The knee motor power to run a half marathon is over 300 W (red arrow), more than two times what we had with the running-optimized design. It wouldn’t be so surprising to need ice packs!

Conversely, visually following the green curve shows that the running-optimized robot wastes more power for walking. Using larger motors sized for running increases the weight of the robot and wastes power when it is standing or walking. The larger motors also pose practical issues like bumping into objects while operating in homes or factories.

Closing Thoughts

Honor’s half-marathon performance was an impressive engineering effort and result. It didn’t need any magical leaps in technology, but the deployment of the capillary motor cooling solution is a notable advance without which this running pace would have been unsustainable. The cooling, weight optimization, and robustness advances may well be useful for more practical purposes like carrying heavy payloads down the line.

A comparison showing two similar humanoid robots, but one has significantly smaller motors on its hips. The Honor Lighting robot [right] has much larger motors driving its legs than the Unitree H1 robot, making it a more efficient runner but a less efficient walker.Left: Wei Zhiyang/Zhejiang Daily Press Group/VCG/Getty Images; Right: VCG/Getty Images

However, the Lightning is not as well-suited to other tasks as a robot designed for greater versatility. Engineering is always characterized by trade-offs, and making the correct ones separates good products from great ones. With consistently improving AI language models, this very human skill is becoming the most valuable one an engineer can have.

The news coverage seemed to overly focus on the fact that the human half-marathon record had been broken by a robot. Machines and humans have very different capabilities and constraints, so why should we ever have expected the half-marathon time for a robot and human to be related? As in Deep Blue’s 1997 defeat of Garry Kasparov in chess, where it couldn’t physically move the pieces, the Honor robot’s capabilities are much narrower than a human running elbow to elbow with other runners while visually navigating the course without GPS. Comparing the robot runner to a human runner is just an apples-to-oranges comparison, which only risks diminishing Honor’s engineering achievement on one hand and human athletic achievement on the other.

Pretrained to Imagine, Fine-Tuned to Act: The Rise of World-Action Models

15 June 2026 at 12:00
Quick glossary for readers new to VLA/WAM terminology VLA Vision-Language-Action model: a robot policy that starts from a pretrained VLM backbone and adapts it...

Quick glossary for readers new to VLA/WAM terminology VLA Vision-Language-Action model: a robot policy that starts from a pretrained VLM backbone and adapts it to generate actions from visual observations and language instructions. Large-scale VLM pretraining is a core part of the recipe. See Pi-0 and GR00T N1. WAM World-Action Model: a policy that starts from a pretrained world-model or video…

Source

Visual Language Models Train Robots to Read Human Emotions

13 June 2026 at 13:00


This article is part of our exclusive IEEE Journal Watch series in partnership with IEEE Xplore.

As robots advance in terms of dexterity and other physical capabilities, it becomes more likely that humans may find themselves working alongside them. If that happens, how will robots’ emotional capabilities need to advance for them to successfully work with people?

In a recent study, researchers trained collaborative robots to read human emotions by not only accounting for facial expressions, but also contextual factors in the interactions as well. Through experiments with 40 volunteers, the researchers then evaluated how a robot’s ability to read human emotions and adjust its behavior in turn impacted a human’s perception of the robot and its capabilities as the two collaborated on tasks. The results—which show that the emotional capabilities of robots only go so far with humans—were published 18 May in IEEE Robotics and Automation Letters.

Seung Chan Hong led the study as part of his undergraduate thesis while studying at Monash University, in Melbourne, Australia. He notes that, while there has been a lot of hype in the advancing physical abilities of robots, this is only one piece of the puzzle. “We need to also innovate when it comes to them actually interacting with humans, not just their physical capabilities,” he says.

This prompted him to dig deeper into the emotional aspects of human-robot interactions. First, Hong and his co-authors decided to train a robot to read human emotions using a vision language model (VLM), which is similar to large language models (LLMs) such as ChatGPT, but which can also take visual inputs.

Training VLMs for Human Emotion Recognition

To evaluate their VLM, which used Gemini 2.5, the researchers had volunteers watch videos of robots handing over objects to humans—with varying degrees of success—and describe the emotions the humans were expressing. Importantly, the volunteers labeling these videos were able to take into account more context in these interactions, rather than reporting solely on the facial expressions of the humans in the video. For example, a person pausing to think with a furrowed brow may simply be concentrating on their task at hand and not necessarily be angry. Contextual factors such as drumming their fingers, pursing their lips, or other behaviors can point to the real cause of a person’s furrowed brow.

The researchers then compared their VLM to a conventional AI system that relies on standard facial analysis and object tracking that is used in human-robot interactions. They found that the VLM outperformed the traditional approach. On a scale from 0 (no similarity in meaning to the emotion identified by the human volunteers) to 1 (a perfect match in meaning), the conventional AI system achieved a score of 0.77. In comparison, the VLM achieved a score of 0.86.

Hong says, “I think [the VLM] was able to align with what human observers were seeing a lot better, because it wasn’t just looking at the person’s face for a brief amount of time, but seeing the whole scene—where the person was and what they were doing, and how they were interacting with the robot.”

In a second experiment, the research team asked 40 volunteers to interact with a robot using their VLM—but purposefully programmed the robot to make an error. The robot then had to offer either an emotionally adaptive apology that accounted for the human’s perceived response to the mistake or a pre-scripted spoken apology.

Participants overwhelmingly preferred the emotionally adaptive response, with 31 out of 40 people favoring this approach over a boilerplate apology.

However, their survey responses underscored how this emotional adaptivity was far less important than the robot’s functionality. After collaborating with a robot that failed in its task, many participants ranked their trust in the robot as lower, regardless of how it apologized for its mistake. “A personalized apology acts as a social lubricant, but it cannot repair the trust lost by the robot failing its physical task,” Hong says.

Interestingly, the VLM classified the emotions of its human partners similarly to human volunteers who observed an interaction from a third-party perspective. But when the VLM’s assessments were measured against humans’ self-reported emotions during the second experiment—the most accurate descriptions of their true emotions—its ability to accurately predict emotions dropped significantly.

“While the VLM is a good observer of outward social cues, it isn’t a mind reader,” Hong says. “It matched third-person human observers well, but it didn’t always align with the users‘ internal, self-reported feelings.”

Together, these results show that robots are not perfect at reading human emotions. So while people might appreciate their efforts, they still ultimately will want competent co-workers.

This story was updated on 15 June 2026 to correct where the research was conducted and clarify that the researchers evaluated the performance of a pre-trained model.

Defining Autonomy for Wellness Robots in Senior Care

11 June 2026 at 10:00


An examination of how socially assistive wellness robots could support the seven dimensions of senior wellness, and how a framework can measure their autonomy.

What Attendees will Learn

  1. Why the senior care crisis exceeds incremental automation. Demographic pressure, workforce shortages, and a daily wellness-programming gap all strain traditional care models.
  2. What defines a wellness robot as a category. The seven ICAA wellness dimensions and eight properties separate these robots from companion and medical devices.
  3. How autonomy can be measured with CRAS. This six-level scale, modeled on the SAEJ3016 driving standard, evaluates four care dimensions.
  4. What maps the road to full autonomy. The paper examines technical capabilities, clinical evidence, and a three-phase roadmap toward the early 2030s.

Deploy Agentic-Ready AI at the Edge with Memory Efficiency in NVIDIA JetPack 7.2

2 June 2026 at 02:00
As AI agents move from the digital world to the physical environment, they can readily use NVIDIA Jetson to accelerate real-world deployment with optimized...

As AI agents move from the digital world to the physical environment, they can readily use NVIDIA Jetson to accelerate real-world deployment with optimized memory and performance. NVIDIA JetPack 7.2 directly supports one-command deployment of NVIDIA NemoClaw, an open source stack that adds privacy and security controls to OpenClaw. It introduces NVIDIA agent skills for Jetson—Jetson device…

Source

How to Post-Train Autonomous Vehicle Models in Closed-Loop with NVIDIA Alpamayo

1 June 2026 at 04:49
Developing autonomous vehicle (AV) policies requires bridging an important gap between training and deployment. Vision-language-action (VLA) models that can...

Developing autonomous vehicle (AV) policies requires bridging an important gap between training and deployment. Vision-language-action (VLA) models that can reason over more complex driving scenes and produce richer intermediate reasoning are predominantly trained in open-loop, where model outputs are directly compared to ground-truth behaviors without considering their effect on the environment.

Source

💾

Develop Physical AI Reasoning, World, and Action Models with NVIDIA Cosmos 3

1 June 2026 at 04:43
Physical AI systems must understand the real world before they can act within it. Robots, autonomous vehicles, and smart spaces need to understand what's...

Physical AI systems must understand the real world before they can act within it. Robots, autonomous vehicles, and smart spaces need to understand what’s happening in their world, predict what’s likely to happen next, and generate actions for specific environments, embodiments, and tasks. NVIDIA Cosmos 3 is a frontier foundation model for physical AI that combines physical reasoning…

Source

❌