We are looking for a QA Engineer to validate agent memory capabilities within an enterprise AI Agent Platform being built from the ground up. The role will focus on ensuring that AI agents reliably retain, retrieve, and manage contextual information across both individual conversations and multiple sessions.
You will test stateful and personalized agent behaviors, memory extraction and retrieval, lifecycle controls, and regression scenarios to ensure that memory improves agent performance without introducing incorrect or irrelevant context.
About the Client
Our client is a large global enterprise operating across multiple markets, with a complex technology landscape and a strong focus on digital transformation and innovation. The organization is actively investing in modern cloud, data, and AI capabilities to enable scalable, secure, and highly automated solutions across its business.
About the Project
The project is focused on building an enterprise-grade AI Agent Platform from the ground up. The platform will provide a standardized foundation for developing, deploying, orchestrating, securing, and observing AI agents across multiple teams and use cases.
The initiative covers the full agent lifecycle and brings together agent orchestration, observability, security, governance, integrations, evaluation, and platform engineering. Engineers joining the project will have an opportunity to influence key architectural and technical decisions and contribute to a new platform rather than maintaining an existing solution.
Skills:
• Agent memory functional testing — short-term memory (in-session continuity), long-term memory (cross-session recall)
• Memory extraction validation — verify Summary and Semantic strategies extract correct facts/preferences
• Memory retrieval testing — validate agents retrieve relevant context, filter irrelevant memories
• Memory lifecycle testing — TTL expiration, user deletion, namespace isolation
• Agent regression testing — verify memory does not degrade agent response quality
• Python test automation for memory workflows
Experience:
• 4+ years QA engineering for AI/ML or conversational AI systems
• Testing stateful agent behaviors (context retention, personalization)
• Functional test automation for API-driven services
• Regression testing methodologies
Nice to have:
• Agent memory framework testing (LangChain, LangGraph)
• Conversational AI evaluation
• AWS AgentCore Memory testing
• GDPR compliance testing (user deletion validation)
At Intellias, where technology takes center stage, people always come before processes. By creating a comfortable atmosphere in our team, we empower individuals to unlock their true potential and achieve extraordinary results. That’s why we offer a range of benefits that support your well-being and charge your professional growth.
We are committed to fostering equity, diversity, and inclusion as an equal opportunity employer. All applicants will be considered for employment without discrimination based on race, color, religion, age, gender, nationality, disability, sexual orientation, gender identity or expression, veteran status, or any other characteristic protected by applicable law.
We welcome and celebrate the uniqueness of every individual. Join Intellias for a career where your perspectives and contributions are vital to our shared success.