Salary
≈ $49k – $125k per year (Estimated)
Location
Remote (Argentina, Brazil)
Seniority
Senior · 2+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Intellectsoft is a global software development company that provides a wide range of services, including mobile app development, web development, cloud computing, data analytics, and more. The company has a team of experienced developers who are committed to delivering high-quality software solutions. Intellectsoft has a strong reputation for its expertise in various technologies, and it has worked with a wide range of clients across different industries.
Our customer's product is an AI-powered platform that helps businesses make better decisions and work more efficiently. It uses advanced analytics and machine learning to analyze large amounts of data and provide useful insights and predictions. The platform is widely used in various industries, including healthcare, to optimize processes, improve customer experiences, and support innovation. It integrates easily with existing systems, making it easier for teams to make quick, data-driven decisions to deliver cutting-edge solutions.
Requirements
- Bachelor’s or Master’s degree in Computer Science or a related field.
- Strong Python coding skills - 7+ years.
- 2+ years of hands-on experience with machine learning and production LLM systems.
- Experience building backend APIs with FastAPI, async patterns, rate limiting, and SQLAlchemy - 3+ years.
- Experience designing maintainable and extensible systems using dependency injection, interfaces, and abstract base classes.
- Experience with vector databases such as Pinecone, Weaviate, or Chroma, as well as hybrid search.
- Strong understanding of RAG architectures, including retrieval, reranking, context assembly, and response generation.
- Hands-on experience with LangChain and LangGraph for building and orchestrating LLM workflows.
- Advanced Python skills, including async/await, type hints, Pydantic, and SOLID principles.
- MLOps experience with MLflow, model versioning, and A/B testing; experience with Langfuse is a plus.
- Experience in NLP and computer vision, including document understanding, OCR, and GPT-4 Vision.
- Experience building feature pipelines, real-time and batch inference systems, and model serving.
- Hands-on experience with Hugging Face is required; experience with LlamaIndex is a plus.
- Familiarity with database technologies such as SQL.
- Good problem-solving skills and the ability to work in a fast-paced, team-oriented environment.
Nice to have skills:
- Understanding of DevOps, CI / CD including: Docker containerization, Azure DevOps pipelines or GitHub Actions, Kubernetes (nice to have);
- Data security including: Multi-tenant data isolation, Secure key management (Azure Key Vault), Audit trail implementation;
- Experience in designing on cloud platform including: Azure (strongly preferred): Azure OpenAI, Blob Storage, Key Vault, Container Registry, AWS or GCP;
- Experience in data engineering in Big Data systems including: Large-scale data processing, ETL/ELT pipelines.
- Rate limiting and quota management for high-throughput API usage.
- Cost management and optimization for LLM usage at scale.
- Document processing expertise (PDF extraction, OCR tooling).
- Production incident management and on-call experience.
- Testing strategies for non-deterministic LLM outputs (e.g., golden datasets, fuzzy matching).
- Domain knowledge in regulated industries (e.g., healthcare/pharma workflows, regulatory compliance) is a plus.
Responsibilities:
- Build, refine, and use ML Engineering platforms and components; develop and implement scalable backend systems, APIs, and microservices using FastAPI.
- Implement MLOps including model KPI measurement, tracking, model drift detection, and model feedback loops.
- Deploy and operationalize ML and Deep Learning models, with a strong focus on LLMs and Generative AI.
- Integrate Azure OpenAI (GPT-4, GPT-4 Vision) and other LLM providers with proper retry logic and error handling.
- Maintain up-to-date knowledge of state-of-the-art technologies such as LLMs, GenAI, and transformer architectures.
- Scale machine learning algorithms to work on massive data sets under strict SLAs.
- Build and orchestrate model pipelines including feature engineering, inferencing, and continuous model training.
- Write backend application code in Python and SQL using strong object-oriented principles and asynchronous programming (asyncio, async/await).
- Implement dependency injection patterns and layered architecture (Service, Foundation, Orchestration, DAL).
- Build LLM observability (e.g., Langfuse) to track prompts, tokens, costs, and latency.
- Develop prompt management systems with versioning and fallback mechanisms.
- Implement Celery (or similar) workflows for asynchronous task processing and complex pipelines.
- Build multi-tenant architectures with client data isolation.
- Implement cost optimization strategies for LLM usage (prompt caching, batch processing, token optimization).
- Integrate third-party APIs and services (e.g., document/OCR services, cloud storage, enterprise systems).
- Collaborate with client-facing teams to understand business context and contribute to technical requirement gathering.
- Write production-ready code that is testable, maintainable, and accounts for edge cases and errors.
- Ensure high quality of deliverables by following architecture/design guidelines, coding best practices, and periodic design/code reviews.
- Write unit tests and higher-level tests to handle expected edge cases and errors gracefully.
- Troubleshoot backend application code using structured logging and distributed tracing.
- Use bug tracking, code review, version control, and other tools to organize and deliver work.
- Participate in scrum calls and agile ceremonies, communicating progress, issues, and dependencies.
- Document application changes and updates, including API documentation via OpenAPI/Swagger.
- Research and evaluate emerging architecture patterns and technologies through rapid learning, proofs-of-concept, and prototypes.
Benefits
- Awesome projects with an impact
- Udemy courses of your choice
- Team-buildings, events, marathons & charity activities to connect and recharge
- Workshops, trainings, expert knowledge-sharing that keep you growing
- Clear career path
- Absence days for work-life balance
- Flexible hours & work setup - work from anywhere and organize your day your way
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
In your city
Automation and AI Developer
8 hours ago
≈ $28k – $116k per year (Estimated) • Remote/Hybrid • Full-Time • Bengaluru • Hyderabad
Python
SQL
Python
FastAPI
Pydantic
Databases
FAISS
OpenSearch
Pinecone
SAP HANA
Snowflake
AI/ML
A2A
AI Agents
AWS Bedrock
AWS Bedrock AgentCore
Claude
Claude Code
Cline
Copilot
Cursor
DeepEval
Embeddings
Gemini
Hallucination
LangChain
LangGraph
Llama
LLM
LLM Guardrails
Model Context Protocol
OCR
OpenAI
Promptfoo
RAG
DevOps
AWS
Azure
GCP
GitHub
Management
Power Automate
Marketing
Salesforce
Apply
Sr. Forward Deployed Engineer
7 hours ago
≈ $62k – $142k per year (Estimated) • In office • Full-Time • 6+ years exp • PhD • Madrid
Apex
JavaScript
Python
TypeScript
Databases
Databricks
Google BigQuery
Snowflake
AI/ML
Agentforce
AI Agents
Claude
Cursor
LangChain
LlamaIndex
LLM
Prompt Engineering
Marketing
Salesforce
Apply
Fullstack-разработчик
7 hours ago
$22k – $38k per year (net) • In office • Full-Time • 3+ years exp • Astana
Python
TypeScript
JavaScript
Python
Alembic
Celery
FastAPI
Litestar
Pydantic
SQLAlchemy
structlog
Databases
ClickHouse
ElasticSearch
MinIO
NATS
Neo4j
PostgreSQL
Redis
AI/ML
LLM
OpenAI
Pydantic AI
Frontend
Mantine
React Hook Form
React.js
Recharts
shadcn/ui
Tailwind CSS
TanStack Router
TanStack Table
Vite
Zustand
Radix UI
DevOps
Amazon S3
Ansible
Docker
Git
GitLab
GitLab CI
Grafana
gRPC
Terraform
QA
Playwright
Pytest
Apply
Applied AI Engineer - Technology Consultant
7 hours ago
≈ $144k – $279k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • New York
Python
AI/ML
Claude
Copilot
Function Calling
LangChain
LLM
Prompt Engineering
PyTorch
RAG
AI Agents
LLM Guardrails
DevOps
AWS
Azure
Robotics
Digital Twin
Apply
Senior .NET Developer (DotNetNuke) (IR-543)
1 month ago
≈ $49k – $111k per year (Estimated) • Remote • Full-Time • Spain
C#
C#
.NET
Apply
Data Architect (IR-541)
1 month ago
≈ $30k – $71k per year (Estimated) • Remote • 6+ years exp • Master's Degree
C#
C#
.NET
Databases
Microsoft Fabric
AI/ML
AI Agents
DevOps
Azure
Analytics
ETL/ELT
Marketing
Marketo
Salesforce
Apply
Embedded AI Engineer (IR-540)
1 month ago
≈ $60k – $104k per year (Estimated) • Remote • Full-Time • Master's Degree
C#
C#
.NET
AI/ML
AI Agents
LLM
Prompt Engineering
Claude
Anthropic
Apply
Data Architect (IR-541)
1 month ago
≈ $66k – $128k per year (Estimated) • Remote • Full-Time • Master's Degree
C#
C#
.NET
Databases
Microsoft Fabric
AI/ML
AI Agents
DevOps
Azure
Analytics
ETL/ELT
Marketing
Marketo
Salesforce
Apply
Forward Deployed Engineer (IR-537)
1 month ago
Remote • Full-Time • 5+ years exp • Master's Degree
Java
Node JS
Python
SQL
C#
JavaScript
C#
.NET
gRPC for .NET
Databases
Snowflake
AI/ML
dbt
Spark
Frontend
GraphQL
DevOps
AWS
Azure
CI/CD
Docker
GCP
gRPC
Kubernetes
Terraform
Apply
This is one of many
368,634 more open roles from verified company boards, updated every day.

