Salvo Software brand banner

AI Developer

Salvo SoftwareBengaluru, IndiaPosted 1h ago
via Workable

About Salvo Software

Salvo Software is a global firm that provides cost-effective software solutions to guide enterprises and startups through digital transformation. With distributed teams across the US, LATAM, and India, we partner with clients to build high-performance, scalable systems that solve complex technical challenges. Our culture values innovation, ownership, and engineering excellence.

Role Overview

We are seeking a highly skilled AI Developer with a strong backend and machine learning engineering background to design, train, optimize, and deploy LLM models in on-prem and offline environments. This role is deeply technical and hands-on, requiring expertise across Python ML stacks, model optimization, local inference frameworks, RAG (Retrieval-Augmented Generation) architectures, MCP (Model Context Protocol) integrations, and DevOps workflows tailored for offline systems.

You will work closely with our engineering and product teams to build end-to-end LLM pipelines — including data preprocessing, supervised fine-tuning, model quantization, evaluation, RAG pipeline design, and deployment using local or air-gapped infrastructure. If you enjoy working with cutting-edge open-source LLMs, building context-aware AI systems, and designing reliable backend pipelines, this role is for you.

Key Responsibilities

Core LLM Development

  • Train and fine-tune LLMs using supervised fine-tuning (SFT).
  • Work with open-source models such as LLaMA, Mistral, Qwen, and similar architectures.
  • Build LoRA / Q-LoRA pipelines for efficient fine-tuning.
  • Implement and optimize data preprocessing workflows, including tokenization and long-context handling.
  • Use and extend Hugging Face Transformers & Datasets for training and inference.
  • Parse and process structured and semi-structured data, including XML/XSD files.
  • Implement document parsing solutions for Office formats (python-docx, OpenXML).

RAG & Context-Aware Systems

  • Design and implement end-to-end Retrieval-Augmented Generation (RAG) pipelines for document-grounded question answering and knowledge retrieval.
  • Build and maintain vector stores and embedding pipelines using tools such as FAISS, Chroma, Weaviate, or pgvector.
  • Optimize retrieval strategies including hybrid search, re-ranking, and chunking approaches tailored for domain-specific corpora.
  • Develop and maintain MCP (Model Context Protocol) server integrations to enable LLMs to interact dynamically with tools, APIs, and external data sources.
  • Design agentic workflows that leverage MCP to give models structured access to internal systems and context in a controlled, auditable manner.

Offline / On-Prem Model Expertise

  • Deploy, run, and maintain models fully offline and in air-gapped environments.
  • Perform model optimization and quantization (GGUF, GPTQ, AWQ, bitsandbytes).
  • Build and maintain inference systems using frameworks like vLLM, TGI, and Ollama.
  • Optimize GPU usage (CUDA, cuDNN, VRAM-aware batching).
  • Maintain local CI/CD pipelines for ML models without cloud dependencies.
  • Manage local model registries, versioning, and artifacts.
  • Ensure RAG and MCP components are fully operational in offline and restricted network environments.

Backend & DevOps

  • Build backend services in Python for ML training and inference workflows.
  • Work with relational databases (Postgres/MySQL) and vector databases for RAG storage layers.
  • Use Docker and Git for reliable development and deployment pipelines.
  • Use Azure DevOps for CI/CD, including local runners when applicable.

Technical Skills

  • Strong experience in Python for backend and ML development.
  • Expertise with ML frameworks such as PyTorch or TensorFlow, scikit-learn, and pandas.
  • Solid knowledge of Postgres or MySQL for data storage.
  • Experience with Docker, Git, and DevOps best practices.
  • Hands-on expertise with LLM training, fine-tuning, and optimization.
  • Experience with Hugging Face Transformers & Datasets.
  • Familiarity with XML/XSD and Office document parsing tools.
  • Experience deploying models with vLLM, TGI, or Ollama.
  • Understanding of quantization techniques (GGUF/GPTQ/AWQ).
  • Experience working with GPU optimization and the CUDA stack.
  • Ability to build solutions for offline, on-prem, and air-gapped environments.
  • Hands-on experience designing and implementing RAG pipelines, including embedding models, vector stores (FAISS, Chroma, Weaviate, or pgvector), and retrieval optimization strategies.
  • Experience building or integrating MCP (Model Context Protocol) servers to connect LLMs with external tools, APIs, and structured data sources.

Nice to Have

  • Experience building agentic systems using MCP in production or near-production environments.
  • Familiarity with advanced RAG techniques such as HyDE, re-ranking, or multi-hop retrieval.
  • Experience managing ML model registries in offline environments.
  • Familiarity with AWS for hybrid deployments.
  • Experience with secure environments, restricted networks, or enterprise compliance requirements.

Soft Skills

  • Strong ownership mindset and problem-solving ability.
  • Ability to work effectively in distributed teams across time zones.
  • Clear communication when discussing complex technical topics with both technical and non-technical stakeholders.

Similar roles

  • Marcura logo

    DA Processing Specialist

    Marcura·Mumbai, India

    DA-Desk process Disbursement Accounts on behalf of many vessel Operators globally, our role in the processing team is to make the job of the Vessel Operator as efficient as possible. We ensure the accuracy of Disbursement Accounts, and we provide insight into costs which Operators would not normally have. Part of this function is the correct and efficient communication with…

    • Hybrid
    • Full-time
  • Marcura logo

    Manager - Billing and Collections

    Marcura·Mumbai, India

    The Manager – Billing and Collections is accountable for the accuracy, integrity and timeliness of the Order-to-Cash (O2C) cycle across ShipServ and Vesselman — from contract capture through invoicing, credit and collections, IFRS 15 revenue recognition support and AR aging, in a multi-entity, multi-currency, PE-backed environment. The role owns monthly revenue closing and reconciliation for both entities, drives O2C process…

    • Full-time
  • Marcura logo

    Global Revenue Assurance & Billing Head

    Marcura·Mumbai, India

    The Head of Global Revenue Assurance & Billing owns end-to-end invoicing, billing, revenue reporting and cash collection across all Marcura Group entities and products — c. 5000 - 8,000 invoices a month supporting. Reporting to the Head of Group Accounting and partnering with the Group CFO, Treasury, and Sales/CSM teams, the role is accountable for the accuracy, integrity and timeliness…

    • Full-time
  • Genetec logo

    Développeur(euse) logiciel Senior/ Senior Software Developer - Clearance

    Genetec·Montréal, Canada

    La dynamique de votre équipe : Clearance, c'est une équipe de développeurs et de testeurs enthousiastes, animés par la curiosité et l’envie de repousser les limites de la technologie. Nous valorisons l'innovation, l'entraide et la créativité pour relever des défis ambitieux et transformer des idées en solutions concrètes afin d’améliorer le quotidien de différents verticaux du secteur publique. Concrètement, nous…

    • Hybrid
    • Full-time
    • Express EntryPR day one, no employer
  • Genetec logo

    Développeur(euse) logiciel / Software Developer - Clearance

    Genetec·Montréal, Canada

    La dynamique de votre équipe : Clearance, c'est une équipe de développeurs et de testeurs enthousiastes, animés par la curiosité et l’envie de repousser les limites de la technologie. Nous valorisons l'innovation, l'entraide et la créativité pour relever des défis ambitieux et transformer des idées en solutions concrètes afin d’améliorer le quotidien de différents verticaux du secteur publique. Concrètement, nous…

    • Hybrid
    • Full-time
    • Express EntryPR day one, no employer
  • Messe Muenchen India logo

    Business Manager-BU1-TLI

    Messe Muenchen India·Mumbai, India

    ROLE SUMMARY We are currently seeking a highly motivated, results-driven, and extensively B2B sales-driven individual to join our Sales and Alliances team as a Business Manager. As the Business Manager for an assigned Project (i.e., a nomenclature for trade fairs), you will be responsible for identifying new business opportunities, building and managing relationships with key clients, and driving revenue growth.…

    • Full-time
  • Bhoomi Group logo

    Front Desk Receptionist

    Bhoomi Group·Mumbai, India

    Company Description Bhoomi Group, established in 1990, is a housing services provider focused on enriching people’s lives by delivering reliable, quality residential solutions. The company is known for its commitment to client satisfaction, innovation, and long-term value creation. Bhoomi Group works closely with clients to understand their housing needs and offers customized services to meet those expectations. Team members join…

    • Full-time
  • TotalEnergies logo

    Key Account Manager West - Heavy Duty (construction & infra)

    TotalEnergies·Mumbai, India

    This role is responsible for growing lubricant sales and managing key customers in the Construction & Infrastructure segment across the West region. Responsibilities Responsible for top line & Bottom line of Customers cluster being handled. To track and maintain accurate market information related Construction & Infrastructure business, Govt Policies. Lead, build & nurture value selling with the customers Maintain top…

    • Full-time