1. Voice edge
AssemblyAI Voice Agent streams English transcripts, spoken replies, and tool calls using short-lived server-issued tokens.
Product & architecture
FixSA Voice is a VALO Systems hackathon concept: a safe, structured front door for reporting everyday public-infrastructure issues.
The problem
Phone queues, fragmented forms, missing landmarks, repeated reports, and opaque follow-up all waste time. Voice can close that gap—if the product preserves consent, structure, safety, and operator control.
Product principles
Architecture
AssemblyAI Voice Agent streams English transcripts, spoken replies, and tool calls using short-lived server-issued tokens.
Zod schemas validate classification, confirmation, duplicate, evidence, status, priority, and SLA operations.
Laravel and MariaDB provide transactional writes, public/private projections, audit history, idempotency, and indexed queues.
The same experience remains demonstrable through synthetic browser-local data when no API is configured.
Privacy by design
Consent appears before recording. Audio is ephemeral in the demo. Public tracking excludes contact, identity, precise-home, transcript, internal-note, and assignment detail.
Limitations
Locations and work orders are synthetic, routing teams are fictional, metrics are simulated, and no record is dispatched to a municipality. Emergency help remains the responsibility of official services.
Roadmap
Accent testing, category vocabulary, field policy, SLA rules, and accessibility research with consenting participants.
Authorised work-order API, identity and abuse controls, geocoding, retention policy, data processing terms, and operator training.
Only expose languages actually supported and tested for the selected model; measure equitable completion and transcription quality.