The problem
A Saudi government client needed an enterprise-grade bilingual (Arabic/English) assistant, with evaluation and monitoring of answer quality.
What I did
- Built the enterprise-grade bilingual (Arabic/English) assistant.
- Added LLM evaluation and observability monitoring focused on response quality.
Outcome
- An enterprise bilingual assistant with built-in evaluation and monitoring of answer quality.
-
Users and apps
- Users (Arabic / English)
-
Orchestration
- Bilingual assistant
-
Models
- LLM
-
Quality
- LLM evaluation
- Observability: answer quality
Connections
- Users (Arabic / English) to Bilingual assistant
- Bilingual assistant to LLM
- LLM evaluation to Bilingual assistant
- Observability: answer quality to Bilingual assistant
Figure 1 Capability view: the public description covers the assistant, its evaluation and its monitoring; implementation details are not public. Diagrams show the components named in public descriptions of the work. They are simplified, not complete system maps.