Containerized Vertical Farming Using Cobots

summary

Video file (mp4)

The gist

Containerized vertical farming (CVF) presents challenges due to space limitations and labor intensity, necessitating automation for key operations like sapling transplantation and harvesting.

In short

The research automated sapling transplantation and harvesting in containerized vertical farms using a collaborative robot (cobot). By using a single human demonstration, the method extracts motion constraints from RGBD images. This allows the robot to generalize its planning for new tasks without needing specific programming for every new tube or object.

Key concepts

Collaborative Robots (Cobots)
These are robots designed to work safely alongside humans in shared workspaces. In this context, they are used in vertical farming containers to perform delicate manipulation tasks like inserting saplings or picking greens, requiring them to follow learned motion constraints.
Segment Anything Model (SAM)
SAM is a deep learning foundation model that excels at identifying and segmenting objects in images. It helps the system quickly recognize and isolate specific parts of the scene, such as the growing tubes or saplings, providing geometric knowledge for planning.
Screw-Geometric Representation
This method describes robot movements not as traditional paths but as sequences of constant screw motions within a mathematical framework called SE(3). This representation is coordinate-invariant and allows the system to capture the underlying physical constraints of a movement, making it easier to generalize across different task setups.
Motion Subgroups in SE(3)
These are specific sets of allowed movements or constraints that define how a robot can move relative to its environment. By extracting these subgroups from a demonstration, the system can create flexible motion plans that satisfy the required physical rules for transplanting or harvesting, even when the exact objects change.

Terminology used across episodes

This episode discusses

The paper

Containerized Vertical Farming Using Cobots · Read on arXiv

Department of Mechanical Engineering, Stony Brook University, USA · Department of Computer Science, Stony Brook University, USA · CubicAcres LLC

Containerized vertical farming is a type of vertical farming practice using hydroponics in which plants are grown in vertical layers within a mobile shipping container. Space limitations within shipping containers make the automation of different farming operations challenging. In this paper, we explore the use of cobots (i.e., collaborative robots) to automate two key farming operations, namely, the transplantation of saplings and the harvesting of grown plants. Our method uses a single demonstration from a farmer to extract the motion constraints associated with the tasks, namely, transplanting and harvesting, and can then generalize to different instances of the same task. For transplantation, the motion constraint arises during insertion of the sapling within the growing tube, whereas for harvesting, it arises during extraction from the growing tube. We present experimental results to show that using RGBD camera images (obtained from an eye-in-hand configuration) and one demonstration for each task, it is feasible to perform transplantation of saplings and harvesting of leafy greens using a cobot, without task-specific programming.

DOI: 10.1109/ICRA57147.2024.10609985

Transcript

Introduction to the show: ident: Robotics Radio. Generated commentary on the latest robotics and control papers.

Rosa: I'm Rosa, and with me are Dev and Taro, guest researcher.

Dev: Today's paper: "Containerized Vertical Farming Using Cobots".

Rosa: Containerized vertical farming (CVF) presents challenges due to space limitations and labor intensity, necessitating automation for key operations like sapling transplantation and harvesting.

Dev: First, who's behind it and why it matters.

Title and authors: Rosa: Well, the title itself tells us they are focusing on using collaborative robots in containerized vertical farming because space is super tight there. Dev It’s interesting that they bring in cobots since traditional mobile manipulators just don't fit into those shipping containers, right? Taro The authors are Mahalingam, Patankar, Phi, Chakraborty, McGann, and Ramakrishnan—they seem to be a solid mix of robotics and autonomy expertise.

Rosa: They're aiming to automate two specific operations: transplanting saplings and harvesting plants. It sounds like they’re trying to solve the labor issue in that environment by automating the physical manipulation steps. Dev And what's exciting about their approach is that they’re not programming a new motion plan for every single plant; they are trying to learn constraints from just one human demonstration.

Taro: That idea of extracting motion constraints from a single human demo to generalize it is where my focus lies, because it suggests a level of adaptability that goes beyond task-specific programming. It implies the system can handle variations in the physical setup without needing entirely new code for every single growing tube configuration.

Rosa: Right, so the core idea is using that demonstration to derive rules for movement, rather than explicitly coding every single insertion or extraction path they need to perform. Dev That moves us away from rigid programming and toward a more learned behavior based on geometric understanding.

The paper's summary: Dev: The summary highlights how they combine a deep learning model, specifically the Segment Anything Model, with geometric knowledge of the tubes and screw-geometric representations of motion into their planning system. Rosa That combination is key because it’s not just relying on vision; it’s using that visual data to define mathematical constraints in SE(three) space <ref:2310.15385#pg0>. Taro So, when the robot needs to transplant something, it uses SAM to figure out where the slot is in three dee, and then those visual features are combined with the demonstration data <ref:2310.15385#pg0>.

Dev: And they represent the demonstration as a sequence of constant screw motions or one-parameter subgroups of SE(three), which they argue is a coordinate-invariant way to describe movement <ref:2310.15385#pg0>. It’s interesting because that representation should theoretically be robust to how you define your starting point in space. Rosa That sounds like a strong theoretical foundation for transferring those constraints from the training demonstration to the actual task instance, which is exactly what they set out to do with different slots.

Taro: I wonder how that mathematical transfer works when the environment changes significantly; if we move outside the exact geometry of the demo, can this constraint transfer still be accurate? It sounds like a major area where you need high autonomy to handle those discrepancies.

Dev: That's a valid concern about robustness. The paper claims this method allows them to define a new sequence of motion subgroup constraints, G′, based on identifying the constant screws that fall inside the sphere around the new objects. It’s like they are extracting a localized rule set for the specific task instance and applying it to plan the next movement.

Rosa: So, in essence, they’ve built a system where one demonstration teaches you *how* to move relative to an object, and then their vision system tells you *where* that object is now so you can apply those learned rules correctly. Taro That dependency on both the visual localization via SAM and the learned motion geometry seems like a clever way to bridge perception and action for this constrained manipulation task.

The paper's improvements: Rosa: The improvements they propose are really about achieving that generalization we talked about earlier, moving past simple task programming. They suggest using the deep learning foundation model, SAM, alongside geometric knowledge of the tubes to define the slot pose estimate in R3. Dev That estimation step is crucial because if the robot doesn't accurately know where the slot is in three dee space, none of that motion constraint transfer will work properly <ref:2310.15385#pg0>.

Taro: I'm interested in how this impacts real-world scenarios where things aren't perfect; for example, what happens when the RGBD data is noisy or if the lighting changes significantly? Does this framework handle those kinds of sensing errors well?

Dev: The experimental validation suggests it's quite resilient, achieving an overall success rate of eighty-three point eight percent in their tests with a Franka Emika Panda manipulator. They showed it could successfully insert saplings into slots with different diameters, like thirty mm and thirty-five mm, while still satisfying those constraints they learned from the demonstration.

Rosa: That's a solid result for handling physical variations in tube sizes, which is exactly what a farming operation needs to do. But what about the harvesting task? Taro Harvesting involves occlusion with foliage, so I wonder if the method can handle that visual ambiguity well when trying to extract those constraints for extraction from the tube.

Dev: For harvesting, they found that even when views were occluded by leaves, the system could still perform it successfully because it used the pose estimates of the planting slots derived from the transplantation task. It seems like reusing prior information helps compensate for temporary visual obstructions.

Conclusion: Rosa: So to wrap up, this paper on "Containerized Vertical Farming Using Cobots" shows a way to use a single human demonstration and deep learning segmentation alongside screw-geometric representations in SE(three) to plan for constrained manipulation tasks <ref:2310.15385#pg0,Containerized Vertical Farming Using Cobots>. Dev The main implication is that we can move toward robots that adapt their motion plans based on learned constraints instead of needing bespoke programming for every single growing tube configuration.

Taro: I think the real impact here is demonstrating how we can create a flexible planning system where the autonomy handles task instance variation by leveraging learned motion subgroups, which opens up possibilities for deploying these systems in less controlled settings.

Rosa: Right, and that’s what makes it so compelling for applications like CVF where labor is scarce and space is limited. It shows a path toward building more versatile robotic systems that can operate without constant manual reprogramming.

Dev: From my end, the success rate of eighty-three point eight percent across different tube specifications gives us a concrete baseline for how reliable this constraint-based transfer method is in practice, provided the initial pose estimation doesn't fail catastrophically.

Taro: I just want to emphasize that while it works well under the tested conditions, future work needs to focus on improving gripper geometry and making the system even more robust against those environmental uncertainties we discussed earlier.

Rosa: Well, it’s clear this research lays a solid foundation for automating repetitive physical tasks in vertical farming using collaborative robots. We’ll keep an eye on how they refine this approach in future iterations of "Containerized Vertical Farming Using Cobots."

More episodes

← Home