Senior AI Engineer - Agentic Systems & Data Pipelines

Remote
Full Time
Experienced

Who We Are

Collaboration.Ai is a mission-focused, AI-powered software and services company based in Minnesota, with employees, partners, and customers around the world. We unite people, technology, and purpose to accelerate breakthroughs that transform industries, empower communities, and create a more sustainable future. We collaborate with organizations across the defense ecosystem, helping them navigate complex challenges and drive transformative change.

Our Products

NetworkOS — NetworkOS is an AI-powered platform that aligns people, purpose, ideas, and expertise in real-time, generating actionable insights to propel movements forward.

CrowdVector — CrowdVector is an integrated solution marketplace and innovation management platform that rapidly uncovers new ideas and advances breakthroughs to fuel movements.

To learn more about us, visit collaboration.ai.

About the Role

You'll build the agentic systems and data pipelines behind NetworkOS's AI capabilities: production agent workflows built on industry-leading agent SDKs and harnesses, MCP servers, and Agent Skills standards; the eval and observability layer that keeps LLM quality measurable; and the ingestion pipelines that turn messy, diverse data sources into queryable knowledge.

This is an execution seat, not an ivory tower. You'll commit code every week, ship agents as product capability rather than demos, and help shape a roadmap that's heading deep into graph + agents territory — for customers in defense, healthcare, and regulated enterprise.

Agents in production. Pipelines that hold. Evals that keep everyone honest.

This opportunity is remote with a preference for candidates in the Twin Cities area (Minneapolis, Saint Paul); however all candidates are encouraged to apply!

What You'll Do

  • Ship production agent systems — design, build, and operate agentic workflows (agent SDKs, MCP servers, Agent Skills standards) powering AI-driven matching, analysis, and data intelligence
  • Operationalize LLM quality — build the eval and observability layer with Langfuse, golden datasets, LLM-as-judge patterns, and FinOps-style tracking so every workflow has measurable quality, cost, and latency
  • Engineer data pipelines — robust ingestion of documents, structured data, and external sources into searchable knowledge bases with quality validation, deduplication, and incremental updates
  • Own retrieval quality — hybrid search combining vector, keyword, and metadata retrieval, continuously improved through reranking, query expansion, and contextual compression
  • Accelerate with AI — build custom MCP tools and Agent Skills that make the whole engineering team measurably faster
  • Execute alongside the team — pair with full-stack engineers on AI integration points, contribute to incident response for AI services, and keep your hands in the code

Our Tech Stack

  • Languages: Python (primary); Kotlin (core platform language at CAI); TypeScript/Node.js and other modern languages (secondary)
  • AI/ML: FastAPI, Pydantic; multi-provider LLM SDKs (Anthropic, OpenAI, and others)
  • Agentic Tooling: Claude Code/Codex/etc.; industry-leading agent SDKs and harnesses; MCP servers; Agent Skills standards
  • LLM Operations: Langfuse + evals (golden datasets, LLM-as-judge); in-house FinOps tracking (token usage, latency, cost); multi-provider orchestration including AWS Bedrock
  • Search & Retrieval: Vector databases, OpenSearch, embedding models
  • Data: PostgreSQL, Amazon S3; streaming pipelines (Kafka/Kinesis) where needed
  • Infrastructure: Docker, Kubernetes (AWS EKS); DataDog + OpenTelemetry observability

What We're Looking For

Must Haves

  • 7+ years of professional software engineering experience, with 3+ years focused on AI/ML or data engineering
  • Production agentic/LLM application experience — built and operated systems around LLM APIs (Anthropic, OpenAI) serving real users: agents, tool-use, or orchestrated LLM workflows
  • Data engineering background — robust, scalable pipelines for AI/ML workloads
  • LLM operations experience — evals and observability for production LLM systems (quality, cost, latency)
  • Production retrieval experience — vector databases and/or search engines (OpenSearch, Elasticsearch)
  • Modern Python stack proficiency — FastAPI, Pydantic, async/await, modern dependency management
  • AI-native workflows — demonstrated ability to leverage Claude Code/Codex or similar agentic coding tools to accelerate development
  • Experience with Docker, Kubernetes, and AWS
  • US citizenship required (DoD contracting — IL4/IL5 environments — and FedRAMP compliance)
Nice-to-Haves
  • Deep agentic ecosystem experience — Agent Skills standards, custom MCP servers, agent SDKs across major vendors
  • Advanced RAG expertise — GraphRAG, agentic RAG, contextual retrieval, reranking strategies
  • Graph data experience — knowledge graphs, graph databases, or graph-based retrieval
  • Model selection & rightsizing — matching models to domain-specific use cases across quality, cost, and latency tradeoffs
  • Streaming data experience (Kafka, Kinesis) for real-time knowledge base updates
  • Research background, open-source contributions, or an advanced degree in ML/IR/NLP

Why Join Collaboration AI?

Real AI engineering, not a wrapper shop. Production agents, hybrid retrieval, continuous evals, and a roadmap heading into graph + agents — with the autonomy to shape how it's built.

AI-native by default. We build with AI, not just for AI. Agentic coding tools (Claude Code/Codex/etc.), agent SDKs and harnesses, MCP servers, and Agent Skills standards are how we work daily — you'll both use and build them.

Work that matters. Defense, healthcare, and regulated industries — SOC 2 and NIST compliance, FedRAMP readiness, and customers whose missions demand AI they can trust.

Small, senior team. Early-stage impact with your work visible from week one. You'll help set the bar for how AI engineering is done here.

Share

Apply for this position

Required*
We've received your resume. Click here to update it.
Attach resume as .pdf, .doc, .docx, .odt, .txt, or .rtf (limit 5MB) or Paste resume

Paste your resume here or Attach resume file

To comply with government Equal Employment Opportunity and/or Affirmative Action reporting regulations, we are requesting (but NOT requiring) that you enter this personal data. This information will not be used in connection with any employment decisions, and will be used solely as permitted by state and federal law. Your voluntary cooperation would be appreciated. Learn more.

Voluntary Self-Identification of Disability
Voluntary Self-Identification of Disability Form CC-305
OMB Control Number 1250-0005
Expires 05/31/2026
Why are you being asked to complete this form?

We are a federal contractor or subcontractor. The law requires us to provide equal employment opportunity to qualified people with disabilities. We have a goal of having at least 7% of our workers as people with disabilities. The law says we must measure our progress towards this goal. To do this, we must ask applicants and employees if they have a disability or have ever had one. People can become disabled, so we need to ask this question at least every five years.

Completing this form is voluntary, and we hope that you will choose to do so. Your answer is confidential. No one who makes hiring decisions will see it. Your decision to complete the form and your answer will not harm you in any way. If you want to learn more about the law or this form, visit the U.S. Department of Labor’s Office of Federal Contract Compliance Programs (OFCCP) website at www.dol.gov/ofccp.

How do you know if you have a disability?

A disability is a condition that substantially limits one or more of your “major life activities.” If you have or have ever had such a condition, you are a person with a disability. Disabilities include, but are not limited to:

  • Alcohol or other substance use disorder (not currently using drugs illegally)
  • Autoimmune disorder, for example, lupus, fibromyalgia, rheumatoid arthritis, HIV/AIDS
  • Blind or low vision
  • Cancer (past or present)
  • Cardiovascular or heart disease
  • Celiac disease
  • Cerebral palsy
  • Deaf or serious difficulty hearing
  • Diabetes
  • Disfigurement, for example, disfigurement caused by burns, wounds, accidents, or congenital disorders
  • Epilepsy or other seizure disorder
  • Gastrointestinal disorders, for example, Crohn's Disease, irritable bowel syndrome
  • Intellectual or developmental disability
  • Mental health conditions, for example, depression, bipolar disorder, anxiety disorder, schizophrenia, PTSD
  • Missing limbs or partially missing limbs
  • Mobility impairment, benefiting from the use of a wheelchair, scooter, walker, leg brace(s) and/or other supports
  • Nervous system condition, for example, migraine headaches, Parkinson’s disease, multiple sclerosis (MS)
  • Neurodivergence, for example, attention-deficit/hyperactivity disorder (ADHD), autism spectrum disorder, dyslexia, dyspraxia, other learning disabilities
  • Partial or complete paralysis (any cause)
  • Pulmonary or respiratory conditions, for example, tuberculosis, asthma, emphysema
  • Short stature (dwarfism)
  • Traumatic brain injury
Please check one of the boxes below:

PUBLIC BURDEN STATEMENT: According to the Paperwork Reduction Act of 1995 no persons are required to respond to a collection of information unless such collection displays a valid OMB control number. This survey should take about 5 minutes to complete.

You must enter your name and date
Human Check*