Isolation
Website-scoped knowledge
Pages, prepared knowledge, questions, and answers remain tied to the active website instead of mixing unrelated customer content.
Technology
RAG + voice readinessGuftara separates website preparation, grounded answering, and speech delivery so each part can be evaluated, secured, and scaled for a real pilot.
System flow
The current MVP completes the website-to-grounded-answer path. Pilot infrastructure will add and qualify multilingual speech delivery.
Read permitted public website pages and preserve a clear record of each scan attempt.
Turn successful pages into website-scoped knowledge that can support later questions.
Find relevant source context and compose an answer within the active website boundary.
Return a reviewable text answer today and qualify multilingual speech delivery during pilots.
Technical foundations
Isolation
Pages, prepared knowledge, questions, and answers remain tied to the active website instead of mixing unrelated customer content.
Quality
The Preview Assistant can surface supporting page titles and URLs so a user can inspect where an answer came from.
Language
The architecture is designed for multilingual retrieval and speech. Exact language and voice support is qualified and published for each pilot.
Scale
Long-running website preparation is separated from interactive questions so each workload can scale independently.
Why cloud credits matter
Credits extend runway while Guftara measures model quality, speech latency, and real usage before committing to long-term capacity.
Runs the product API, account workflows, website coordination, and background preparation jobs.
Stores website records, Indexed Pages, scan state, and vector-capable representations used to ground answers.
Supports embeddings, answer generation, evaluation, and controlled experiments with managed or open models.
GPU capacity is needed to evaluate low-latency multilingual text-to-speech models for pilot conversations.
A focused pilot helps us measure grounding quality, language needs, response latency, and the infrastructure required for production.