Research questionHow can co-speech gesture generation preserve semantic grounding and speech alignment without sacrificing biomechanical smoothness?Co-speech gesture systems must express lexical meaning while timing movements to speech. Semantic gestures and rhythmic beat gestures can compete, producing weak grounding, poor alignment, or jittery, physically implausible motion.