Tiếp theoHướng dẫn tiếp theo
Cách dịch menu và biển hiệu bằng camera điện thoại
AI trực quan
HƯỚNG DẪN AI trực quan
AI form-checking apps use pose estimation: a computer vision model finds points like your shoulders, hips, knees and ankles in each video frame and measures joint angles to judge things like squat depth, tempo and symmetry.
They give useful, low-cost feedback on obvious errors. But a single camera misses depth, spine position and bracing, so pain, rehab or heavy lifting still call for a qualified coach or physical therapist.
Pose estimation is the core technology. A neural network looks at each video frame and outputs coordinates for body landmarks, called keypoints. OpenPose, released by Carnegie Mellon researchers in 2017, made real-time multi-person pose estimation widely available. Google's MediaPipe Pose tracks 33 landmarks, and MoveNet predicts 17 keypoints in the common COCO format. Apple's Vision framework also includes body pose detection. Form apps build on models like these. They connect the keypoints into a stick figure, calculate angles at the knee, hip and elbow, and compare them against thresholds, such as whether the hip dropped below the knee. This works well for things a camera can see clearly: squat depth from the side, rep counts, tempo, left-right differences, and big errors like half-range push-ups. It works poorly for others. A single camera produces a flat image, so movement toward or away from the lens is hard to measure. Knees caving inward are nearly invisible from the side, while depth is hard to judge from the front. Small changes in spine position, bracing, breathing, pressure through the feet and grip are mostly invisible. Bar path needs separate object tracking. Loose clothing, a squat rack or plates blocking the view, poor lighting and fast movements all make keypoints less accurate. The main misconception is that an app saying 'good form' means the lift is safe. The thresholds are generic, and people's bodies differ. A lifter with long thigh bones may need more forward lean to squat well, and ankle mobility changes what works. The app does not know your injury history or how heavy the weight feels. For better results, set the camera on a stable surface at about hip height, keep your whole body in frame, and film from the side and the front. See a coach when you feel pain, when you are learning heavy barbell lifts, during rehab, or when you stop making progress.
Visual AI có thể tự động hóa các nhiệm vụ kiểm tra, phát hiện và gắn thẻ trên quy mô lớn.
Các nhóm sáng tạo có thể tạo nguyên mẫu nhanh hơn với ít sửa đổi thủ công hơn.
Các hoạt động có thể sử dụng tín hiệu hình ảnh và video mà trước đây khó xử lý.
Pose models keep getting more accurate and efficient enough to run on phones, and combining video with other data, such as depth sensors or wearable motion sensors, may help with what a single camera cannot see. Better evidence on whether app feedback actually reduces injuries or improves technique is still needed. It is reasonable to expect apps to handle more exercises and to explain their feedback more clearly. It is less reasonable to expect them to replace a trained professional who can watch your movement from every angle, ask about pain and adjust the plan.
A home lifter films squats from the side with the phone at hip height. The app reports that their hips stop above knee level on most reps, so they work on reaching depth with lighter weight.
A lifter's app says their squats look fine, but when they film from the front their knees clearly cave inward. That is a problem the side view could not show.
A beginner uses an app's rep counter during push-ups and notices it misses reps when their loose hoodie hides their elbows. They switch to a fitted top and better lighting.
Someone returning from a back injury uses an app to count reps between physical therapy visits. They let the therapist, not the app, decide when to add weight.
Quyền và sự đồng ý về hình ảnh có thể trở thành rủi ro pháp lý nếu nguồn gốc xuất xứ không rõ ràng.
Hiệu suất của mô hình có thể khác nhau tùy theo ánh sáng, nhân khẩu học và môi trường.
Kết quả dương tính giả có thể không được chú ý trừ khi ngưỡng tin cậy được theo dõi.
Xác định tiêu chí chấp nhận về độ chính xác, thu hồi và chi phí lỗi.
Kiểm tra với dữ liệu phù hợp với điều kiện sản xuất thực tế.
Thêm đánh giá của con người đối với những dự đoán có độ tin cậy thấp hoặc tác động cao.
Theo dõi sự trôi dạt của mô hình và xác nhận lại sau khi thay đổi máy ảnh hoặc tập dữ liệu.
Free newsletter
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
AI form-checking apps use pose estimation: a computer vision model finds points like your shoulders, hips, knees and ankles in each video frame and measures joint angles to judge things like squat depth, tempo and symmetry. They give useful, low-cost feedback on obvious errors. But a single camera misses depth, spine position and bracing, so pain, rehab or heavy lifting still call for a qualified coach or physical therapist.
Pose models find keypoints such as shoulders, hips, knees and ankles. Apps then connect them and measure angles.
MediaPipe Pose tracks 33 landmarks. MoveNet predicts 17 keypoints in the COCO format.
A single camera gives a flat image, so movement toward or away from the lens is hard to see. A front view is needed.
Internal factors like bracing, breathing and foot pressure, along with small spine position changes, cannot be seen reliably in video keypoints.
Generic thresholds cannot account for individual bodies, injury history or how heavy the weight feels.
Tiếp tục học hỏi
Đã chọn thêm hướng dẫn cho chủ đề này
Tiếp theoHướng dẫn tiếp theo
Cách dịch menu và biển hiệu bằng camera điện thoại
AI trực quan