AI Tutor - Audio [Remote]
$90k - $200kxAI
- Remote job
About xAI
xAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All engineers are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.
About the Role
As an AI Tutor specialized in audio, you will contribute to xAI's mission by training and refining Grok to excel in voice interactions, sound processing, and auditory experiences. Key to this role is possessing an exceptional vocal quality (a great voice), exceptional audio perception (a great ear), hands-on background in sound manipulation through editing software, hosting shows, or composing tracks, and demonstrating advanced proficiency in audio principles, techniques, and quality standards.
Responsibilities
You will use proprietary software to provide labels, annotations, and inputs on projects involving sound clips, voice recordings, and auditory elements. You must support the delivery of high-quality curated data that ensures clear, professional audio output and accurate representation of sonic details. In this role, you will collaborate with technical staff to develop tasks that improve AI's ability to handle voice modulation, noise reduction, and immersive sound design. You’ll also work with technical staff to improve annotation tools for efficient workflows.
Required Qualifications
- Demonstrated high proficiency in Audio Engineering, Music Production, Communications, or a related field.
- Outstanding vocal presence suitable for recording and demonstration purposes.
- Proven skills in audio post-production, episode creation, or musical arrangement, with a deep grasp of acoustics, mixing, and signal processing.
- Proficiency in evaluating and enhancing sound quality, with strong expertise in tools for waveform analysis, equalization, and format compatibility.
- Strong ability to reference professional standards, equipment, and best practices for annotating and refining auditory content.
- Strong communication, interpersonal, technical, and detail-oriented skills.
- Commitment to developing AI that masters sophisticated audio capabilities.
Preferred Qualifications
- Portfolio of audio work, such as edited tracks, hosted episodes, or produced compositions, shared on platforms or professionally.
- Experience in sound design, broadcasting, studio recording, or roles involving auditory critique and optimization.
Location & Other Expectations
- This position is based in Palo Alto, CA, or fully remote.
- The Palo Alto option is an in-office role requiring 5 days per week; remote positions require strong self-motivation.
- If you are based in the US, please note we are unable to hire in the states of Wyoming and Illinois at this time.
- We are unable to provide visa sponsorship.
- Team members are expected to work from 9:00am - 5:30pm PST for the first two weeks of training and 9:00am - 5:30pm in their own timezone thereafter.
- For those who will be working from a personal device, please note your computer must be a Chromebook, Mac with MacOS 11.0 or later, or Windows 10 or later.
- You must own and have reliable access to a smartphone.
Compensation
$45/hour - $100/hour
Benefits:
Hourly pay is just one part of our total rewards package at xAI. Specific benefits vary by country, depending on your country of residence you may have access to medical benefits. We do not offer benefits for part-time roles.
xAI is an equal opportunity employer.
$35 - $45 per hour
xAI is seeking an AI Tutor specialized in multilingual audio to enhance Grok's voice interactions. The role involves curating and annotating audio data to improve speech recognition globally. Candidates should have native proficiency in Italian and be proficient in English...AudioHourly payRemote workFlexible hours$45 - $100 per hour
A leading AI research group is seeking an AI Tutor - Audio Specialist in Palo Alto, CA. This full-time role involves refining AI models through high-quality audio datasets. Candidates should have expertise in audio engineering and experience in sound design. Ideal applicants...AudioHourly payFull timeRemote work$90k - $200k
Role Description Mercor is partnering with a leading AI research organization to engage professionals with advanced expertise in insurance... ..., or machine learning workflows Comfort with recording short audio or video sessions for data training purposes Work Environment &...AudioFull timeWork at officeLocal areaRemote work$45 - $100 per hour
...Description Mercor is partnering with a leading AI research group to engage mathematics... ...refining cutting‑edge AI systems. AI Tutor – Applied Math Specialist will play a key... ...and video formats, including annotations, audio recordings, or video sessions. Use...AudioHourly payContract workWork at officeLocal areaRemote work$45 - $100 per hour
Role Description Mercor is partnering with a leading AI research group to engage audio professionals in a high‑impact project focused on advancing AI... ...auditory and voice interaction capabilities. As an AI Tutor - Audio Specialist, you will play a key role in refining...AudioHourly payFull timeContract workWork at officeLocal areaRemote work$200k - $240k
...Archetype AI is hiring a Senior AI Research Scientist as a backfill for Jamie. You will join the core research team and work on machine... ...sensor data or other non-image data. ~ Experience in robotics, audio, or another physical-world ML domain. ~ A track record of...AudioWork at office2 days per week3 days per week- ...A growing AI technology startup is seeking an ML Engineer to design and deploy production-grade systems. The role involves using Python and collaborating with teams to optimize customer interactions through advanced AI applications. Candidates should have a degree in...Audio
$197k - $291k
# Software Engineer Manager II, Audio and Video, YouTubeGoogle • onsite • Mountain View, CA, USA • full\_timePay: USD 197000.00 - USD 29... ...achieve as a team with sellers, shape the future of advertising in the AI-era, and make a real impact on the millions of companies and...AudioTemporary work- Google DeepMind is seeking a Product Manager for Applied AI to bridge AI research and real-world enterprise impact, focusing on image, video, and audio applications in media and entertainment. You will partner with clients, Research Engineers, and cross-functional teams...Audio
- Google DeepMind is seeking a Product Manager for its Applied AI team in Mountain View to guide products from conception to launch, bridging... ...and deliver frontier AI solutions, including image, video, and audio applications for media and entertainment. You will define product...Audio
- ...focus is on combining time series / sensor data, language, vision, audio, and other real-world signals into unified models that can... ...deep research challenges, practical deployment impact, and the chance to contribute to a fast-emerging area of AI. #J-18808-Ljbffr...Audio
- A gaming technology company in California seeks an AI Trainer to enhance multimodal models by annotating text and improving dialogue interactions. Ideal candidates are fluent in conversational English, have a passion for gaming, and thrive in fast-paced environments. The...AudioHourly payFlexible hours
$188k - $274k
...system and component-level specifications for performance, power, and thermals. Support critical user journeys, including on-device AI for audio and ecosystem interoperability across Google’s platforms.Evaluate the critical trade-offs between power, cost, and form factor....Audio- ...machine learning, as well as a strong educational background, preferably with a PhD. You will lead research projects that integrate AI technologies into Google products, making a significant impact on the future of conversational AI. The role offers a competitive salary...Audio
$180k
xAI in Palo Alto is seeking a Multimodal Engineer to develop advanced AI experiences across image, video, and audio modalities. You will enhance multimodal capabilities, improve data quality, and design evaluation frameworks for cutting-edge AI systems. The ideal candidate...Audio- Google DeepMind in Mountain View, CA seeks a Research Scientist for Gemini Audio i18n to advance multilingual audio models. You will design large-scale experiments, prototype architectures, and evaluate A2A speech systems across languages. You will publish findings and...Audio
$180k
A leading AI firm in Palo Alto is seeking a Backend Engineer for the Grok Voice Product team. You will design and build low-latency voice... ...will have proficiency in Python or Rust, experience in voice/audio AI, and a strong engineering background. The position offers a competitive...Audio$207k - $300k
...seeks a Research Scientist to set up large-scale tests, deploy ideas, and drive progress across machine learning, NLP, and data-driven AI research. The role involves publishing findings and collaborating with universities worldwide, with a focus on real-world AI problems...AudioWorldwide- A digital gaming solutions company is seeking an AI Trainer to improve multimodal models. In this role, you will annotate text transcripts related to game design and work on enhancing user experience through data accuracy. Candidates should be fluent in English, have a...AudioFlexible hours
- ...Scientist to design and deploy large-scale experiments in speech, audio, and multilingual models. You will prototype architectures, run... ...A2A approaches, cross-language support, and delivering practical AI advances across products, while contributing to the research community...Audio
- ...building hardware; electronics systems and semiconductors where AI can design and create beyond human cognitive limits. About the... ...working with multi-modal models (e.g., combining text, image, or audio inputs). Bonus Points Background in competitive...Audio
$147k - $210k
...programming languages, or 1 year of experience with an advanced degree.1 year of experience with one or more of the following: Speech/audio (e.g., technology duplicating and responding to the human voice), reinforcement learning (e.g., sequential decision making), ML...Audio$147k - $210k
...programming languages, or 1 year of experience with an advanced degree.1 year of experience with one or more of the following: speech/audio (e.g., technology duplicating and responding to the human voice), reinforcement learning (e.g., sequential decision making), ML...Audio- ...superhuman multimodal intelligence. You will work across vision, audio, video, and text, building data pipelines, training infrastructure... ...reasoning, world modeling, tool use, and interactive human‑AI collaboration at web/petabyte scale. #J-18808-Ljbffr SpaceXAIAudio
- ...in the world. They are used to power the largest consumer-facing AI applications available, across categories like health, fitness, learning... ...QualificationsExperience with voice AI, TTS, STT, or real-time audio and speech systemsBackground in gaming, CCaaS, or media and...AudioFull timeContract workWork at officeRelocation
$174k - $252k
...Processing, or equivalent practical experience.3 years of experience in deep learning research and development, including generative AI, audio and video synthesis, diffusion models, and autoregressive generative models.One or more scientific publications in venues like...Audio$300k - $333k
...partnerships) to ensure firm engagement foundations.Analyze the enterprise AI landscape, market trends, and engaged offerings to identify... ...(fine-tuning, agent harnesses), generative image, video, and audio models, and technical software development to build products and...Audio$207k - $300k
...develop, and deploy novel multimodal conversational agents.Develop audio-first models capable of orchestrating and planning complex... ...and evaluation of ML models.Experience building or implementing AI/ML-driven features or infrastructure (e.g., working with Large Language...Audio$159k - $230k
Collaborate on groundbreaking features like AI agentic task completion, helping screen-reader users seamlessly accomplish digital chores while carefully managing AI verbosity and audio feedback.Design and iterate on critical multi-modal features like TalkBack, the TalkBack...AudioWorldwideShift work$226.15k - $252k
...advanced capabilities, such as large-scale streaming and specialized audio logic within the orchestration framework.Guarantee the quality of... ...inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing...AudioFull timeTemporary workWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Tutor - Audio [Remote]. Be the first to apply!



