Join Sycomp, a global IT services and logistics provider, as a remote AI Engineer. You will design and build production-grade AI pipelines and autonomous agent systems on AWS, handling messy, unstructured, and multi-modal data. This role focuses on model selection, data engineering, evaluation frameworks, and clear collaboration with stakeholders.
Responsibilities
Design and build AI pipelines for long-form media, unstructured documents, streaming data, and multi-modal inputs, including chunking, context management, retry logic, and output validation.
Architect autonomous agent systems with tool selection, planning loops, memory strategies, evaluation, and guardrails.
Build on AWS using services like Bedrock, S3, Lambda, Step Functions, SageMaker, OpenSearch, and ECS/EKS.
Make pragmatic decisions on model selection, prompt design, fine-tuning vs RAG vs agentic workflows, and cost/latency/quality tradeoffs.
Own the data layer: ingestion, transformation, storage, indexing, and surfacing data to models.
Establish evaluation frameworks and observability to track system improvements.
Collaborate with engineers and stakeholders to translate ambiguous business problems into technical designs and shipped systems.
Requirements
4+ years of software engineering experience, with at least 2 years building AI/ML systems in production.
Mixpanel is an AI-powered digital analytics platform for product teams that helps organizations track user behavior, measure conversions, and improve retention. The company provides a unified platform combining product and web analytics, experimentation, and feature flags.
Hands-on AWS experience, including Bedrock or equivalent foundation model services, and at least a few of S3, Lambda, Step Functions, SageMaker, OpenSearch, DynamoDB, ECS/EKS.
Demonstrable experience building AI pipelines or agentic systems that handle non-trivial data shapes such as long documents, audio/video, large structured datasets, or streaming inputs.
Strong intuition for prompt engineering, context window management, chunking strategies, and output validation.
Proficiency in Python and at least one LLM framework such as LangChain, LlamaIndex, LangGraph, Bedrock Agents, or Strands.
Solid data engineering fundamentals: cleaning, transforming, embedding, indexing, and retrieving data.
Experience designing evaluation pipelines for non-deterministic systems.
Nice to Have
Experience with Azure (Azure OpenAI, AI Foundry, AI Search) or GCP (Vertex AI, Gemini APIs).
Background in multi-modal systems such as vision-language models, audio transcription pipelines, or video understanding.
Experience with vector databases like Pinecone, Weaviate, pgvector, or OpenSearch k-NN, and hybrid retrieval strategies.
Familiarity with fine-tuning, distillation, or model adaptation techniques.
Experience with MLOps tooling, model versioning, and A/B testing for AI systems.
Contributions to open source AI/ML projects.
Benefits
Health insurancedental insurancevision insuranceremote work