flex-piapache-2.0
RoboTwin 2.0 3D — T5 Text Embedding Cache
Precomputed UMT5-XXL text embeddings for 1,039,891 unique task prompts from the RoboTwin 2.0 3D dataset, cached to avoid expensive GPU recomputation during model training.
Downloads3K
Episodes1039891
Why This Matters for Physical AI
This embedding cache accelerates training of multimodal vision-language models for robotics by precomputing expensive text encodings, enabling efficient scaling of video-instruction datasets for embodied AI systems.
Technical Profile
- Modalities
- language
- Episodes
- 1039891
- Data Format
- safetensors
- Annotation Types
- language_instructions
- License
- apache-2.0
Community Signals
Top 25% by downloads
Access
Need custom language data?
Claru builds purpose-built datasets for any environment applications with dense human annotations and quality assurance.
Request a Sample Pack