AI Engineer FastAPI & AWS
Bangalore,
hybrid |
Full-time | Design and scale production-grade AI-powered REST APIs with FastAPI, Pydantic, and AWS (Bedrock, EC2, EMR, Lambda)
|
Location
|
Bangalore, India (Hybrid)
|
|
Level
|
Anyone between Mid to Senior level
|
|
Employment
type
|
Full-time, permanent
|
|
Practice
|
REST APIs with FastAPI, Pydantic, and AWS (Bedrock, EC2, EMR, Lambda)
|
About Power
Tech Consulting
Power Tech Consulting (PTC) is an Australian data, AI, and digital-transformation consultancy. We help government and enterprise clients turn complex data and technology challenges into secure, reliable, and high-performing outcomes, built on precision, accountability, and partnership. Our team spans enterprise architecture, data engineering, generative and agentic AI, cloud, and modern business applications.
Role Overview
As an AI Engineer (FastAPI & AWS) at Power Tech Consulting, you design, build, and scale production-grade REST APIs that expose AI and LLM capabilities to enterprise and government clients. You own the full API lifecycle from schema design and endpoint implementation to deployment and scaling on AWS working alongside data engineers, architects, and our applied-AI research team to deliver fast, reliable, well-documented services.
Key Responsibilities
Design and build RESTful AI APIs in FastAPI, implementing full CRUD (GET, POST, PUT, PATCH, DELETE) resources.
Define request/response schemas, validation, and serialization using Pydantic models.
Integrate AWS Bedrock foundation models and other LLM services behind clean, versioned API contracts.
Design async, non-blocking endpoints for high-throughput and streaming AI workloads.
Deploy and scale API services on AWS EC2 and Lambda, including auto-scaling and load-balancing configuration.
Build and orchestrate data-processing workflows on AWS EMR to support AI pipelines.
Implement authentication, authorization, rate-limiting, and API versioning (e.g. OAuth2, JWT, API keys).
Write automated tests (unit, integration, load) and maintain OpenAPI/Swagger documentation.
Implement caching, connection pooling, and performance tuning for scalable, low-latency APIs.
Set up CI/CD, logging, monitoring, and observability for API services in production.
Required Skills & Experience
Strong Python and backend software-engineering fundamentals.
Hands-on production experience building REST APIs with FastAPI, including full CRUD endpoint design.
Strong command of Pydantic for data validation, schema modelling, and serialization.
Hands-on experience with AWS Bedrock for generative AI / LLM integration.
Practical experience with AWS EC2, Lambda, and EMR for hosting and scaling API and data workloads.
Solid understanding of async Python (asyncio, async/await) for high-concurrency API design.
Experience with API security, authentication/authorization, and rate-limiting patterns.
Experience with cloud-native deployment, containerisation (Docker), and infrastructure-as-code.
Understanding of API scalability patterns: load balancing, caching, horizontal scaling, and queuing.
Desirable
Experience with GenAI frameworks (LangChain, LangGraph, CrewAI or similar) and RAG architectures.
Familiarity with vector databases (Pinecone, FAISS, pgvector).
API gateway experience (AWS API Gateway) and event-driven architectures (SQS, SNS, EventBridge).
MLOps / LLMOps tooling and model monitoring.
AWS certifications (AWS Certified AI Practitioner, Solutions Architect, or similar).
Responsible-AI / AI-governance experience.
What We Offer
Challenging, high-impact work across enterprise and government clients.
A collaborative team of experts and clear pathways for growth.
Support for professional certifications and continuous learning.
Flexible, hybrid working.
Competitive remuneration aligned to your experience and level.