AI/ML Software Engineer

Location
Plano
Workplace
On-site
Compensation
$65 – $70

About this role

Role: AI/ML Software Engineer

Location: Plano, TX
Duration: 6 Months Contract
Employment: W2 (USC/GC Only)
Pay Rate: $65-70/hr w2
Interview: 2 rounds virtual; final round face-to-face

Job Overview

We are looking for a Senior Full Stack & Agentic AI Engineer to design, develop, and deploy production-grade AI applications and intelligent agent solutions.

The ideal candidate will have strong software engineering and full-stack development experience, along with hands-on expertise in Generative AI, Agentic AI, RAG, LLMs, and conversational AI.

Candidates with 7+ years of software development experience are encouraged to apply, provided their background strongly aligns with Full Stack and Agentic AI development.

Key Responsibilities

  • Design, develop, and deploy scalable Full Stack and Agentic AI applications.
  • Build AI agents, multi-agent workflows, RAG pipelines, and LLM-powered applications.
  • Develop backend services and APIs using Java and/or Python.
  • Build modern frontend applications and integrate them with AI/ML services.
  • Implement RAG solutions using chunking, embeddings, vector databases, and retrieval techniques.
  • Work with AWS Bedrock and Anthropic Claude models and APIs.
  • Use AI frameworks such as LangChain, LlamaIndex, Semantic Kernel, or CrewAI.
  • Develop conversational AI solutions, including text chatbots and voice-enabled applications.
  • Apply prompt engineering techniques including system prompts, few-shot prompting, tool calling, and structured outputs.
  • Build scalable microservices and event-driven architectures using AWS services.
  • Implement CI/CD, Docker, Kubernetes/EKS/ECS, and Infrastructure as Code.
  • Develop and integrate REST and GraphQL APIs.
  • Implement LLM evaluation, observability, guardrails, content filtering, and responsible AI practices.
  • Troubleshoot, optimize, and support production AI applications.

Required Qualifications

  • 7+ years of hands-on software development experience.
  • Strong proficiency in Java and/or Python.
  • Hands-on production experience with Generative AI, LLMs, and Agentic AI.
  • Strong experience with RAG architectures, embeddings, chunking, and vector databases.
  • Experience with one or more of:
    • LangChain
    • LlamaIndex
    • Semantic Kernel
    • CrewAI
  • Hands-on experience with AWS Bedrock and Anthropic Claude.
  • Strong understanding of prompt engineering and LLM application development.
  • Experience with REST APIs, microservices, and event-driven systems.
  • Experience with AWS services including:
    • Lambda
    • Step Functions
    • API Gateway
    • S3
    • DynamoDB
    • SQS
  • Experience with Docker, Kubernetes/EKS/ECS, CI/CD, and Infrastructure as Code.
  • Experience building conversational AI, chatbots, or voice-enabled AI applications.


Requirements

Preferred Qualifications

  • Experience with Pinecone, OpenSearch, pgvector, or FAISS.
  • Experience with LLM evaluation and observability frameworks.
  • Experience implementing AI guardrails, content filtering, and responsible AI practices.
  • Experience integrating speech-to-text and text-to-speech technologies.
  • Financial services or banking industry experience.
  • Experience working in enterprise-scale and highly regulated environments.

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

Know someone who'd be great for this?