@HuggingApps: text-to-motion that can follow precise commands and sequences "a person walks forward, then sits down on the floor" — a…
Summary
PRISM is a 1.4 billion parameter text-to-motion model that can generate precise motion sequences from text commands, returning SMPL-X parameters for use in 3D rigs.
View Cached Full Text
Cached at: 08/24/26, 04:01 PM
text-to-motion that can follow precise commands and sequences
“a person walks forward, then sits down on the floor” — and it does them in that order 🚶🪑
PRISM (1.4B) hands you back SMPL-X params, so you can drop it on your own rig
▶️ on Spaces https://t.co/sXprTxZ2y0 https://t.co/ofyysWqpKI
PRISM Text-to-Motion - a Hugging Face Space by hugging-apps
Source: https://huggingface.co/spaces/hugging-apps/prism-text-to-motion Fetching metadata from the HF Docker repository...
Similar Articles
PRISM: A Benchmark for Programmatic Spatial-Temporal Reasoning
PRISM is a large-scale benchmark of 10,372 human-calibrated instruction-code pairs for evaluating programmatic video generation, with a funnel-style framework of four metrics. Evaluation of seven LLMs reveals a significant gap between code executability and spatial coherence.
PRISM: Prosody-Integrated Multi-Agent Reasoning Framework for Empathetic Spoken Dialogue
PRISM is a multi-agent framework that decouples speech perception, response generation, and speech synthesis to improve empathetic spoken dialogue by integrating prosodic cues with LLM reasoning and external knowledge tools.
PRISM: Perception Reasoning Interleaved for Sequential Decision Making
This paper introduces PRISM, a framework that integrates Vision-Language Models and Large Language Models through a dynamic question-answering pipeline to improve sequential decision-making in embodied AI tasks.
PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI
PRISM is a closed-loop framework that treats prompt engineering as a continuous reliability problem for enterprise conversational AI. It automates test generation, simulation, evaluation, and repair, achieving 99% reliability and reducing authoring time from days to minutes.
@oliviscusAI: this tool can track perfect 3D motion. rtmlib is a lightweight pose estimation library covering full body, hands, face,…
rtmlib is a lightweight, open-source pose estimation library that supports full-body, hand, face, and animal pose tracking, built on rtmpose and vitpose models, with a built-in Gradio web UI.