The three pillars of Hawkeye-1
Multimodal perception
Hawkeye-1 sees and listens in parallel. It combines what it observes on camera with what it hears in speech to form a unified, multimodal understanding of the user. This is what makes it a true perception system, not just a computer-vision model.Environmental awareness
Context doesn’t stop at the face. Hawkeye-1 continuously analyses the video feed, tracking environmental changes, detecting key gestures and behaviours, and triggering relevant actions as needed. Whether a user holds up a product for a visual question, steps away from the screen, or changes their surroundings, Hawkeye notices — and the agent adapts.Emotional intelligence
Hawkeye-1 reads body language, vocal tone, and subtle facial expressions to render emotional understanding the way humans do — with real-time awareness of context and conversational flow. The result is an agent that doesn’t just react to words, but responds to meaning.Developer-friendly by design
Hawkeye-1 is built for integration. A single API flag enables vision perception, and natural-language prompts let you configure what to track — objects, gestures, on-screen activity, or all of the above.- Enable Hawkeye with a single flag — flip one parameter and your agent gains real-time scene analysis, effortlessly and precisely.
- Answer visual questions — the model sees and understands the user’s environment, responding to visual queries about objects, scenes, and actions just like a human would.
- Custom action triggers — automate responses and fire tool calls at the right moment based on what Hawkeye perceives (timing, tone, gesture, or environment).
What you can build with Hawkeye-1
Healthcare
Understand patient mood and engagement in real time. Trigger tool calls at the right moment based on timing, tone, and conversational flow.
Education & training
Avatars that gauge learner engagement, adjust pacing, and provide personalised guidance — making training more effective at scale.
Customer service
Handle FAQs and guide users through complex flows while reading frustration or confusion before it escalates.
Sales
Qualify leads and showcase products with avatars that read buyer signals — hesitation, interest, objection — and adapt in real time.
Recruiting
Screen candidates with adaptive interviews that evaluate answers, communication style, confidence, and engagement.
Better together: Hawkeye-1 + Huma
Hawkeye-1 is powerful on its own, but reaches its full potential when paired with the Huma family of avatar renderers. Together they deliver:- Visual realism + situational awareness — Huma renders photorealistic faces; Hawkeye-1 adds real-time perception of who’s watching.
- Emotionally aware + contextually responsive — Huma expresses emotion through micro-expressions; Hawkeye-1 tells it exactly when and how.
- Real-time emotional sync — expression, rhythm, and context synchronise into a single seamless conversational experience.
Key characteristics
Next steps
Vision Understanding
How to enable Hawkeye-1 on your agent and configure what to track.
Models overview
See how Hawkeye-1 fits alongside the Huma family of renderers.
