Jobs Companies Salvo Software AI Developer

Über diese AI Developer Stelle bei Salvo Software

Salvo Software · Vor Ort · Bengaluru, Karnataka, India

About Salvo Software

Salvo Software is a global firm that provides cost-effective software solutions to guide enterprises and startups through digital transformation. With distributed teams across the US, LATAM, and India, we partner with clients to build high-performance, scalable systems that solve complex technical challenges. Our culture values innovation, ownership, and engineering excellence.

Role Overview

We are seeking a highly skilled AI Developer with a strong backend and machine learning engineering background to design, train, optimize, and deploy LLM models in on-prem and offline environments. This role is deeply technical and hands-on, requiring expertise across Python ML stacks, model optimization, local inference frameworks, RAG (Retrieval-Augmented Generation) architectures, MCP (Model Context Protocol) integrations, and DevOps workflows tailored for offline systems.

You will work closely with our engineering and product teams to build end-to-end LLM pipelines — including data preprocessing, supervised fine-tuning, model quantization, evaluation, RAG pipeline design, and deployment using local or air-gapped infrastructure. If you enjoy working with cutting-edge open-source LLMs, building context-aware AI systems, and designing reliable backend pipelines, this role is for you.

Key Responsibilities

Core LLM Development

  • Train and fine-tune LLMs using supervised fine-tuning (SFT).
  • Work with open-source models such as LLaMA, Mistral, Qwen, and similar architectures.
  • Build LoRA / Q-LoRA pipelines for efficient fine-tuning.
  • Implement and optimize data preprocessing workflows, including tokenization and long-context handling.
  • Use and extend Hugging Face Transformers & Datasets for training and inference.
  • Parse and process structured and semi-structured data, including XML/XSD files.
  • Implement document parsing solutions for Office formats (python-docx, OpenXML).

RAG & Context-Aware Systems

  • Design and implement end-to-end Retrieval-Augmented Generation (RAG) pipelines for document-grounded question answering and knowledge retrieval.
  • Build and maintain vector stores and embedding pipelines using tools such as FAISS, Chroma, Weaviate, or pgvector.
  • Optimize retrieval strategies including hybrid search, re-ranking, and chunking approaches tailored for domain-specific corpora.
  • Develop and maintain MCP (Model Context Protocol) server integrations to enable LLMs to interact dynamically with tools, APIs, and external data sources.
  • Design agentic workflows that leverage MCP to give models structured access to internal systems and context in a controlled, auditable manner.

Offline / On-Prem Model Expertise

  • Deploy, run, and maintain models fully offline and in air-gapped environments.
  • Perform model optimization and quantization (GGUF, GPTQ, AWQ, bitsandbytes).
  • Build and maintain inference systems using frameworks like vLLM, TGI, and Ollama.
  • Optimize GPU usage (CUDA, cuDNN, VRAM-aware batching).
  • Maintain local CI/CD pipelines for ML models without cloud dependencies.
  • Manage local model registries, versioning, and artifacts.
  • Ensure RAG and MCP components are fully operational in offline and restricted network environments.

Backend & DevOps

  • Build backend services in Python for ML training and inference workflows.
  • Work with relational databases (Postgres/MySQL) and vector databases for RAG storage layers.
  • Use Docker and Git for reliable development and deployment pipelines.
  • Use Azure DevOps for CI/CD, including local runners when applicable.

Requirements

Technical Skills

  • Strong experience in Python for backend and ML development.
  • Expertise with ML frameworks such as PyTorch or TensorFlow, scikit-learn, and pandas.
  • Solid knowledge of Postgres or MySQL for data storage.
  • Experience with Docker, Git, and DevOps best practices.
  • Hands-on expertise with LLM training, fine-tuning, and optimization.
  • Experience with Hugging Face Transformers & Datasets.
  • Familiarity with XML/XSD and Office document parsing tools.
  • Experience deploying models with vLLM, TGI, or Ollama.
  • Understanding of quantization techniques (GGUF/GPTQ/AWQ).
  • Experience working with GPU optimization and the CUDA stack.
  • Ability to build solutions for offline, on-prem, and air-gapped environments.
  • Hands-on experience designing and implementing RAG pipelines, including embedding models, vector stores (FAISS, Chroma, Weaviate, or pgvector), and retrieval optimization strategies.
  • Experience building or integrating MCP (Model Context Protocol) servers to connect LLMs with external tools, APIs, and structured data sources.

Nice to Have

  • Experience building agentic systems using MCP in production or near-production environments.
  • Familiarity with advanced RAG techniques such as HyDE, re-ranking, or multi-hop retrieval.
  • Experience managing ML model registries in offline environments.
  • Familiarity with AWS for hybrid deployments.
  • Experience with secure environments, restricted networks, or enterprise compliance requirements.

Soft Skills

  • Strong ownership mindset and problem-solving ability.
  • Ability to work effectively in distributed teams across time zones.
  • Clear communication when discussing complex technical topics with both technical and non-technical stakeholders.
Bereit, sich bei Salvo Software zu bewerben?
Bei Salvo Software bewerben

Ähnliche Jobs

Salvo Software
AI Developer
Salvo Software
⚡ Früh bewerben Mexico · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 2 Wo.
Salvo Software
Field Sales Representative
Salvo Software
⚡ Früh bewerben Portland, Oregon, United State... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 3 Wo.
Salvo Software
AI Developer
Salvo Software
⚡ Früh bewerben United States · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 3 Wo.
Salvo Software
Administrative & Personal Assistant
Salvo Software
⚡ Früh bewerben Mexico City, Mexico City, Mexi... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 3 Wo.
Salvo Software
Scan Tool Field Technician
Salvo Software
⚡ Früh bewerben Rigby, Idaho, United States Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Mon.
ProArch
Platform Engineering - Intern
ProArch
⚡ Früh bewerben Bengaluru, Karnataka, India · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 10 Std.
Quanticate
Senior Programmer III
Quanticate
⚡ Früh bewerben Bengaluru, Karnataka, India Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 11 Std.
Rystad Energy
MIT Campus Placement Drive 2026
Rystad Energy
⚡ Früh bewerben Bengaluru, Karnataka, India Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 11 Std.
Fivetran
Technical Program Manager
Fivetran
⚡ Früh bewerben Bengaluru, Karnataka, India, A... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 15 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Salvo Software

Alle Jobs bei Salvo Software ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos