Description
At Network Solutions, we’ve been trusted for decades to help people get online and stay ahead.
We’ve been here since the beginning of the internet, and we’re still building for what comes next.
As the original digital identity authority, we help secure domain names, protect brands, and safeguard the infrastructure businesses rely on.
We empower our customers to own and manage the assets that define them online, while delivering enterprise-grade security to protect against virtual threats.
Our team leverages modern, AI-accelerated tools to streamline how businesses manage their digital presence, making the most of our decades of experience.
The Network Solutions team is here to help online businesses protect what’s theirs and build for tomorrow.
That’s why millions trust us to protect their domains, brands, and websites every day.
The impact you’ll make as a Lead AI Back-End Engineer
• Set the technical architecture for our agent powered platform, choosing the right mix of LLMs, retrieval technologies, agent frameworks, and microservice patterns to meet reliability, cost, quality, and latency targets.
• Drive architecture for high throughput APIs built with .NET, C#, Python 3.11+, FastAPI async, SQLModel, and Semantic Kernel, from design documents through production rollout.
• Lead multi agent orchestration across handoff, sequential, parallel, and supervisor patterns, combining knowledge grounded and tool calling agents across OpenAI GPT 5.6 Sol, Terra and Luna, Google Gemini 3.7 Flash, Anthropic Claude Sonnet 5 and Opus 5, xAI Grok 4.6, and future model providers.
• Guide teams implementing Retrieval Augmented Generation on Azure AI Search, pgvector, Chroma, and equivalent vector and hybrid retrieval platforms, ensuring index quality, filtering, evaluation, and safety controls.
• Own end to end CI/CD pipelines using Bitbucket Pipelines, Jenkins, or similar platforms, including linting, type checking, security scanning, testing, containerization, and deployment.
• Mentor and hire engineers and adopt AI coding agents such as Cursor and Claude Code as engineering force multipliers while maintaining strong code quality, security, and review standards.
• Champion observability and FinOps for LLM workloads using structured JSON logging, OpenTelemetry tracing, Langfuse, evaluation frameworks, and cost dashboards, keeping latency and cost per request within defined SLOs.
• Partner with Product and Security to translate business goals, compliance requirements, emerging AI capabilities, and user feedback into a pragmatic technical roadmap.
Must have experience
• 7+ years building and scaling production backend systems, including 2+ years in a technical lead, staff, principal, or equivalent role.
• Expert in REST API design and development using Python FastAPI or .NET C# APIs, with strong experience in dependency injection, middleware, profiling, performance optimization, and distributed systems.
• Hands on leadership experience with Semantic Kernel or equivalent agent and LLM frameworks, including production agent orchestration, tool calling, structured outputs, and multi step workflows.
• Delivered at least one production RAG system or pipeline using a vector or hybrid search platform such as Azure AI Search, pgvector, Chroma, or equivalent, with measurable latency and quality KPIs.
• Deep PostgreSQL expertise plus SQLModel, SQLAlchemy 2, and Alembic migrations at scale.
• Proven track record integrating multiple frontier and cost optimized model providers, such as OpenAI GPT 5.6, Gemini 3.x, Claude 5, Grok 4.x, or equivalent, including model routing, structured output, reasoning, tool calling, and fallback strategies.
• Fluency with Poetry, Docker, GitHub Actions, Azure DevOps, Jenkins, or Argo, along with infrastructure as code fundamentals and blue green or canary release strategies.
• Strong people leadership through code reviews, architectural guidance, technical mentoring, roadmap planning, hiring, and cross team communication.
Nice to have
• Experience with message queue and event driven architectures using RabbitMQ, Kafka, Azure Service Bus, or equivalent.
• Experience with GPU inference fleets, self hosted open weight models, or serverless model hosting.
• Experience with model gateways, intelligent model routing, MCP, agent interoperability, prompt caching, and LLM cost optimization.
• Familiarity with automatic evaluation pipelines, agent evaluations, LLM as judge techniques, safety guardrails, and cost aware prompt and context engineering.
Why join the Newfold AI team?
You’ll steer the core intelligence behind our AI products, building an LLM agnostic, agent driven platform that serves millions of users while meeting enterprise grade standards for reliability, security, governance, and economics.
If you thrive on significant technical challenges, enjoy mentoring strong engineers, and want the opportunity to shape both architecture and engineering culture, we’d love to meet you.
Why you’ll love us.
Work-life balance.
Our work is thrilling and meaningful, but we know balance is key to living well.
We celebrate one another’s differences.
We’re proud of our culture of diversity and inclusion.
We foster a culture of belonging.
Our company and customers benefit when employees bring their authentic selves to work.
We have programs that bring us together on important issues and provide learning and development opportunities for all employees.
We have 20 + affinity groups where you can network and connect with Newfolders globally.
We care about you.
We are a family, and we care about you and your family’s physical and mental health by providing competitive HMO benefits – 175k MBL with one free dependent upon one year of service! We also give out Punctuality Bonus, Generous Vacation policy, and much more! Where can we take you? We’re fans of helping our employees learn different aspects of the business, be challenged with new tasks, be mentored, and grow their careers.
Unfold new possibilities with #teamnewfold!
#NetworkSolutions