We use cookies. Find out more about it here. By continuing to browse this site you are agreeing to our use of cookies.
#alert
Back to search results
New

Linguist II

Spectraforce Technologies
500 West Peace Street (Show on map)
Aug 04, 2026
Job Title: Linguist II

Location: US - Remote (PST time preferred)

Duration: 12 months (Possibility of Extension or Conversion)

Job Description

Linguist II - Summary

We are looking for a linguist to help develop language components for AI-powered products, including large language models (LLMs) and voice-enabled technologies. We are seeking candidates with solid linguistic data analysis skills, programming familiarity, and language technology experience to contribute to data collection, synthetic data generation, and annotation tasks in support of LLM/AI training, evaluation, alignment, and AI agent development.

Must-Have Skills

  1. Native or near-native fluency in English and at least one additional language.
  2. Knowledge of syntax, semantics, pragmatics, sociolinguistics, corpus linguistics, and other areas of linguistics.
  3. Familiarity with Large Language Models (LLMs), their applications and data practices (training data, evaluation, prompting, fine-tuning).
  4. Exposure to LLM evaluation methodologies (human evaluation, automated metrics, adversarial testing).
  5. Experience working with semantic ontologies, taxonomies, or intent/slot frameworks.
  6. Proficiency using AI Agents/Chatbots.
  7. Experience with database queries and data analysis processes (SQL, Python, spreadsheets, R, Unix, or others).
  8. Experience working with speech and text data in multiple languages.
  9. Comfortable working in a fast-paced, highly collaborative environment with evolving priorities.



Nice-to-Have Skills

  1. Master's degree in Linguistics, Computational Linguistics, Language Technologies, or a related field.
  2. Familiarity with machine learning frameworks, NLP libraries, and tools (e.g., Hugging Face, spaCy, NLTK, PyTorch).
  3. Exposure to statistical language modeling or training data pipelines.
  4. Strong organizational skills and attention to detail.



Years of Overall Experience Required?

  • 2+ years of experience in Linguistics, Language Technologies, NLP, or AI/ML data operations (or equivalent)



Degrees/Certifications Required?

  • Bachelor's degree in Linguistics, Computational Linguistics, Computer Science, Speech Science, or related field



Job Responsibilities

  • Apply linguistic expertise in syntax, semantics, pragmatics, and sociolinguistics to support LLM and generative AI systems.
  • Collaborate with linguists, data operations teams, and ML engineers on data collection, curation, annotation, and localization efforts for model training and fine-tuning.
  • Contribute to the development and maintenance of annotation schemas and guidelines for LLM training data (e.g., instruction-tuning, preference labeling, RLHF).
  • Evaluate and quality-check datasets used for pre-training, fine-tuning, and alignment of language models.
  • Support the development of programmatic methods for generating synthetic annotated data at scale.
  • Assist in model evaluation efforts including prompt-based testing, red-teaming, and linguistic error analysis.
  • Contribute to AI safety and responsible AI practices through linguistic review of model outputs (e.g., hallucination detection, bias identification, tone and pragmatic appropriateness).
  • Participate in experiments to assess data quality, annotation consistency, and downstream model performance.



Surrounding Team & Key Projects

  • help develop language components for AI-powered products, including large language models (LLMs) and voice-enabled technologies



How Will Performance Be Measured?

Meeting deadlines, executing tasks with accuracy, etc.

Interview Process

How Many Rounds of Interviews? 2 rounds

Types of Interviews:

  • 1st round General interview
  • 2nd round Technical interview (python,SQL)


Interview Duration: 45 min each
Applied = 0

(web-77cf7d65c7-jdxdg)