No applications. Teams find and message you directly.
✓ Direct apply: you message the hiring person. No ATS.
We are looking for a Middle+ ML Engineer with experience in building ML pipelines and proficiency in Python and PyTorch. The role involves data preparation, model fine-tuning, and integration with AI services.
We are looking for a Middle+ ML Engineer with experience in building ML pipelines and proficiency in Python and PyTorch. The role involves data preparation, model fine-tuning, and integration with AI services.
Responsibilities
- Build data preparation pipelines: collection, cleaning, normalization, deduplication, and annotation
- Form training, validation, and independent evaluation samples, control leaks and overlaps
- Create baselines for extraction, classification, entity and relationship identification tasks
- Fine-tune models, prepare instructional sets, train adapters, and compare results with baseline models
- Define metrics, stratify samples, and analyze errors by document types, scenarios, and fields
- Work with entity and relationship extraction to build and update graph representations of data
- Deploy local inference and optimize it for available infrastructure
- Work with tensor parallelism, batching, context length, GPU memory, and quantization
- Measure latency, throughput, utilization, and document processing costs
- Version data, models, training configurations, and experiment results
- Participate in integrating models with applied AI services and production pipelines
Requirements
- Candidates of middle+ level and above are considered
- Practical experience in building ML pipelines used in production
- Proficient in Python, PyTorch, and Hugging Face Transformers
- Practical experience in fine-tuning language models, data preparation, and working with PEFT
- Ability to build accurate quality assessments and control leaks between training and test data
- Experience with local inference of LLM and understanding of GPU memory architecture
- Knowledge of trade-offs between quality, speed, context length, batching, and quantization
- Familiarity with Linux, CUDA, SQL, Docker, and Kubernetes
- Ability to go beyond notebooks and bring model solutions to stable operation
✓ Direct apply: you message the hiring person. No ATS.