Skip to content
N
New York Technology Partners

Machine Learning Engineer

Industries
United StatesMid-Senior levelContractor

About this role

Role: AI/ML with Generative AI

Location: Malvern, PA/ Charlotte, NC- HYBRID

Duration : Long Term Contact

Core skills:

  • Strong Python skills including libraries like LangChain, LangIndex, PyTorch, SageMaker SDK, psycopg2
  • Experience with Docker and AWS ECR
  • Strong AWS experience with with the following services:
  • Bedrock
  • SageMaker
  • IAM
  • Glue
  • S3
  • Lambda
  • CodeCommit and CodePipeline
  • Experience creating quick apps using Streamlit, NodeJS, or some other app framework
  • Education and or experience in developing NLP models for text classification, completion, summarization, generation
  • Experience using embeddings models to create vector embeddings and working with vector databases
  • Understanding of RAG architecture, retrieval optimization, and tradeoffs of splitting methods
  • Familiarity with benchmarks for model evaluation and methods of determining vector similarity
  • Experience with scaling ML training workloads using distributed training techniques on GPU and/or developing microservices for AI/ML/GenAI products
  • Data Preprocessing and Analysis: Work with large-scale datasets, preprocess the data, and perform in-depth analysis to derive meaningful insights, patterns, and trends for AI model training.

Preferred candidates will have:

  • Real world experience fine-tuning models, methods of fine-tuning, and data preprocessing for fine-tuning
  • AWS Solutions Architect and or AWS Machine Learning Specialty certifications

Explore related Data & AI jobs

Compare this role with all Data & AI jobs and open roles at New York Technology Partners.

You can also browse all Data & AI jobs to find similar openings by category, seniority, remote setup, and location.

Other jobs at New York Technology Partners

New York Technology Partners has no open Data & AI positions right now.

Browse all jobs
© Dataaxy. All rights reserved.Job data is gathered from publicly available sources or contributed by users.