SRE LLM Engineer
Saransh Inc
Role: SRE ML focus
Location - Sunnyvale/Austin
Mandatory Employer Information for Every Submission
Vendors must provide the following employer details:
- Full Company Name
- Company Revenue
- Total Employee Headcount
- Complete Business Address (as per USCIS records)
- Owner Name (as per USCIS documentation)
- Owner Email Address
- Owner Mobile Number
- Owner's LinkedIn Profile
JOB Details:
Responsibilities
- Design and implement cloud solutions, build MLOps on cloud (AWS or GCP)
- Build CI/CD pipelines orchestration by GitLab CI, GitHub Actions, Flux, Kustomize, Circle CI, Airflow or similar tools
- Data science model containerization, deployment using docker, VLLM, Kubernetes
- Data science model review, run the code refactoring and optimization, containerization, deployment, versioning, and monitoring of its quality
- Data science models testing, validation and tests automation
- Communicate with a team of data scientists, data engineers and architects, document the processes
- Develop and deploy scalable tools and services for our clients to handle machine learning training and inference
Qualifications:
- 6+ years of experience in ML Ops with strong knowledge in Kubernetes, Python, MongoDB and AWS.
- Good understanding of Apache SOLR.
- Proficient with Linux administration.
- Knowledge of ML models and LLM.
- Ability to understand tools used by data scientists and experience with software development and test automation
- Ability to design and implement cloud solutions and ability to build MLOps pipelines on cloud solutions (AWS or GCP)
- Experience working with cloud computing and database systems
- Experience building custom integrations between cloud-based systems using APIs
- Experience developing and maintaining ML systems built with open-source tools
- Experience with MLOps Frameworks like Kubeflow, MLFlow, DataRobot, Airflow etc., experience with Docker and Kubernetes
- Experience developing containers and Kubernetes in cloud computing environments
- Familiarity with one or more data-oriented workflow orchestration frameworks (Kubeflow, Airflow, Argo, etc.)
- Ability to translate business needs to technical requirements
- Strong understanding of software testing, benchmarking, and continuous integration
- Exposure to machine learning methodology and best practices
- Good communication skills and ability to work in a team
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the SRE LLM Engineer in Austin, TX vacancy
$152k - $241.5k
...time large language model inference? Join NVIDIA’s TensorRT Edge-LLM team and help shape the next generation of edge AI for... ...equivalent experience in Computer Science, Electrical/Computer Engineering, or a closely related field.4+ years of relevant software development...SuggestedFull time- ...intend for the selected candidate for this role to work on site in the specified location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and continuously improve Schwab's enterprise streaming platform ecosystem. This role combines deep...SuggestedFull timeWork at office
$152k - $195k
...Senior Site Reliability Engineer Austin, TX (Hybrid) SecurityScorecard is the global leader... ...Required Qualifications ~6+ years in SRE, DevOps, or Infrastructure roles, with... ...experience. ~ Hands-on experience integrating AI/LLM tooling into engineering or operational...Suggested$121.4k - $218.6k
...challenges? Join our critical AI Hardware SRE Team! The AI Hardware SRE team is... ...breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing... .... Leveraging advanced AI utilities and LLM-assisted development paradigms where appropriate...SuggestedWork experience placementWork at office- A competitive technology firm located in Austin, TX is looking for an AI Engineer to enhance their search platform by integrating semantic search and LLM techniques. The ideal candidate will have 10+ years of experience in AI/ML with a strong background in ElasticSearch...Suggested
- ...at Schwab. We are an integrated product, engineering, strategy and risk team, all based in San... ...seamlessly with AI Engineering teams to integrate SRE practices early in the development... ...e.g., Gemini, Claude, OpenAI), deploying LLM-powered applications to production and...Full time
- ...ICON is seeking a pragmatic AI Software Engineer to design and ship AI-powered tools, workflows, and autonomous agents for its Government Technology team. The role is remote with periodic travel to ICON HQ in Austin, TX. You will own end-to-end AI feature development...Remote work
- ***Must be submitted with LinkedIn Profile*** SRE Engineer Must-haves: Candidate should have sound knowledge in Observability & Monitoring (Splunk, Grafana, Hubble) Candidate should be proficient in Scrum Facilitation & Agile Delivery Candidate should have proficiency...
- ...SRE Engineer Experienced SRE Engineer with experience in designing, managing and supporting distributed systems across multi-cloud environments. This position is Hybrid on Austin, Texas. Candidates need to be locals, or be open to relocate. What You'll Do:...Local areaRelocation
- ...Job Description Job Description SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities on this team. Location: Remote | Time Zone: (US, Canada, Brazil, Chile, Colombia,...Remote work
- ...Responsibilities Escalation points for junior site reliability engineers during complex or high-impact incidents. Manage and execute... ...with deployment of AI infrastructure to include clustered GPUs, LLM deployment and maintenance, and understanding of model integration...Temporary workWork experience placementFlexible hoursNight shift
- ...SRE DevOps EngineerTekfortune is a fast-growing consulting firm specialized in permanent, contract & project-based staffing services... ...experts can help you find the best job for you. Role: SRE DevOps Engineer Location: Austin TX Duration: 6 months Required Skills:...Permanent employmentContract workRemote work
$98.58k - $138.02k
Austin, TX/Akron, Ohio/Irvine, CA / SF Bay Area/Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized...Full timeWork at office- ...3 days a week)Position SummaryWe are seeking a Site Reliability Engineer to ensure the high level of service and operation excellence for... .... This product requires the establishment of a product specific SRE team.Essential Functions Automation & Infrastructure as Code: Design...Full timeLocal area3 days per week
- ...work on site in the specified location(s).As a Senior Reliability Engineer, you will help shape the reliability, scalability, and... ...operations teams, you will apply Site Reliability Engineering (SRE) principles to improve system availability, accelerate delivery,...Full timeWork at office
- ...to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the... ...Engineering, IT, or equivalent practical experience.5+ years in SRE, DevOps, or software development roles.Strong experience with Microsoft...Temporary workCasual workWorldwide
- ...encouraged to come as they are and do their best work. The TeamThe 2K SRE team owns the infrastructure behind every player connection—All... ...multiple clouds and regions while partnering with network engineers, systems architects, and game studio developers. This is an ownership...
$152k - $241.5k
...next wave of artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and help build the next generation... ...programming languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,...Full time- ...Dimensional leverages the rapidly evolving state of the art to engineer scalable, innovative, and research driven solutions to improve our... ...for our business. About the Role:We are looking for a Senior SRE to serve as the operations owner for the developer tooling ecosystem...Full timeLocal area
$165k - $241.4k
...Collaboration, and Observability portfolios Your ImpactThe FedRAMP SRE team is focused on our Federal region’s platform. The team is... ...efficient, functional and very effective.We’re looking for talented engineers with a software or operations background, experienced in...Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,... ...platform infrastructure serving millions of users. As a Senior SRE, you will be a strong technical contributor who implements best...Work at officeLocal area
- Job Description:About the Role: We are looking for a Senior SRE to join our Platform Engineering team as the operations owner of our observability platforms. You’ll be responsible for the reliability, scalability, and continued evolution of the tools that give our engineering...Full time
- Selby Jennings is seeking a Senior Site Reliability Engineer to scale and support critical workflow orchestration and automation platforms across the organization. The role sits in Platform Engineering, delivering highly available and resilient infrastructure for business...
- ...or more bounded contexts of the NeoCloud SRE platform — the multi-region substrate that... ...monitor. Alert, Correlation & SLO: alert-engine-framework, alert-correlation, slo-... ...agent-codegen, agent-sandbox, per-Region LLM gateway. Global SRE Management: maintenance...Full timeContract workLocal area
$110.7k - $171.8k
Tink is looking for a SQL Server Database Engineer in Austin, Texas, focusing on production operations and reliability engineering (SRE). This hands-on role requires expertise in T-SQL and PowerShell, emphasizing production support and performance tuning. You will be responsible...- ...AI, and you are the future of Salesforce.Lead Account Solution Engineer - Public Sector Do you want to be part of an outstanding team that... ...to Artificial Intelligence, specifically generative AI and LLM conceptsProven oral, written, presentation, and interpersonal communication...Full timeLocal areaRemote work
- Excelon Solutions in Austin, TX seeks a Site Reliability Engineer / Database Reliability Engineer to design, deploy, and maintain CockroachDB clusters across cloud and on-prem environments. You will monitor health, performance, capacity, and security; implement backup,...
- ...scale. We move fast, we ship often, and we believe the best engineers care as much about the product they're enabling as the systems and... ...leads, who know when deterministic logic or statistics beat an LLM and vice versa, and who care about the customer experience as much...Full timeContract workShift work
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking...Work at officeLocal areaRemote workWorldwideFlexible hours- Visa is seeking a SQL Server Database Engineer in Austin, Texas. This role focuses on the reliability and performance of SQL Server databases and requires strong T-SQL and PowerShell skills. As part of a hybrid team, you'll ensure database availability, troubleshoot production...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE LLM Engineer. Be the first to apply!
Related searches
- site reliability engineer remote Austin, TX
- site reliability engineer sre Austin, TX
- site reliability engineer Austin, TX
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer



