Voice Capture
User speaks a messy natural-language command through wearable glasses.
Voice-first automation platform on Meta Ray-Ban glasses — one spoken command triggers multi-step workflows across 12+ services, including Gmail, Slack, Notion, and Fetch.ai Agentverse. Built as BarelyAtWork at LA Hacks 2026, in a team of four over 36 hours.
"send yesterday's demo clips to the team and block 30 minutes tomorrow morning"
{
"workflow": "share_demo",
"trigger": "voice",
"steps": [
{ "app": "gmail",
"action": "send",
"params_valid": true },
{ "app": "slack",
"action": "post",
"params_valid": true },
{ "app": "calendar",
"action": "block",
"params_valid": true }
],
"status": "ready"
}User speaks a messy natural-language command through wearable glasses.
Gemini extracts the workflow goal, target apps, actions, and required parameters.
A constrained schema checks missing or invalid fields, then repairs the workflow before execution.
The final JSON triggers coordinated actions across email, messaging, calendar, and productivity tools.
Reliable workflows require predictable schemas, not open-ended model output.
Voice commands are incomplete by nature, so the system fills and corrects parameters before running.