A world-leading technology solutions provider, renowned for delivering cutting-edge innovation and strategic insights, is seeking a talented professional to join their dynamic team.
The Role
- Design and implement high-quality, customer-specific synthetic data and RAG / knowledge pipelines.
- Configure, ground, demonstrate, and validate CCAI, voice, and chat deployments without real customer PII.
- Design generation and ingestion pipelines to load data into Google Cloud Platform and AWS services.
- Analyze customer domains, including intents, entities, knowledge topics, and languages.
- Generate synthetic conversation transcripts for voice and chat, along with supporting content like customer/agent profiles and knowledge-base articles.
- Partner with Conversational Platform Specialists and DevOps to integrate data and automate pipelines.
What You'll Need
- 4+ years of experience in data engineering, conversation design operations, applied NLP data work, or knowledge-pipeline engineering.
- Working knowledge of how conversational platforms consume training, FAQ, transcript, and retrieval-grounded knowledge data.
- Strong judgment regarding synthetic-data quality, retrieval quality, and privacy safety.
- Experience with LLM-assisted synthetic data generation in a production or implementation setting is a plus.
- Familiarity with BigQuery, S3, and document stores as knowledge sources is preferred.
- Multilingual data generation or evaluation experience is a bonus.
What's On Offer
- Opportunity to work on innovative conversational AI projects.
- Collaborate with a diverse and experienced team.
- Contribute to cutting-edge solutions in a rapidly evolving field.
Apply via Haystack today!