Ruby Labs

AI Engineer

Ruby Labs GE, FR
Full-time Posted 3 months ago

Role overview

We are looking for a Senior AI Engineer (Node.js / Next.js / TypeScript) to join our team and help advance our AI infrastructure. You'll work within a modern tech stack, focusing on model performance, reliability, and cost efficiency.

You'll take ownership of prompt systems, structured outputs, and LLM workflows built on LangChain or LlamaIndex. The role also covers observability and evaluation using Langfuse and AI gateways such as OpenRouter, with the goal of consistently improving model quality and operational efficiency. You'll drive key AI features from early experimentation all the way through to production.

Responsibilities

  • check_circle Advanced Prompt Engineering: Designing complex, dynamic prompt templates with conditional logic and efficiently reusing information and context within prompts to maximize generation quality and reasoning.
  • check_circle Structured Outputs & Schemas: Implementing various response schemes (JSON mode, function calling, Zod/JSON schemas) to ensure AI outputs are predictable and ready for seamless integration into application logic.
  • check_circle Prompt Engineering & Evaluations: Building robust evaluation pipelines and using Langfuse to collect feedback and score the quality of responses in real time.
  • check_circle Tracing & Debugging: Performing deep debugging of complex LLM chains using Langfuse traces to identify bottlenecks and optimize for cost, latency, and context window usage.
  • check_circle AI A/B Testing: Running systematic experiments across different models via OpenRouter (e.g., comparing Claude 3.5 Sonnet vs. GPT-4o) and analyzing results based on quantitative metrics.
  • check_circle Data-Driven Decisions: Making deployment decisions for new prompts or models strictly based on quantitative benchmarks and trace data, rather than intuition.
  • check_circle Output Scoring & Analysis: Developing scoring systems to analyze the “Problem Solution” chain and identify root causes of hallucinations or logic errors using Langfuse analytics.
  • check_circle Model Performance & Fine-Tuning: Regularly re-evaluating model performance as new architectures emerge and performing fine-tuning when necessary to meet specific domain requirements.
  • check_circle Node.js & Next.js: Deep knowledge of the stack to build reliable services and handle complex LLM-generated data.
  • check_circle Dynamic Prompting Skills: Proven experience in building prompts where content is highly dependent on input variables and context injection.
  • check_circle OpenRouter Experience: Experience working with unified APIs, managing rate limits, and selecting the most cost-effective models for specific tasks.
  • check_circle Langfuse (or similar): Understanding of LLM observability principles — setting up tracing, creating test datasets, and integrating scoring systems.
  • check_circle Evaluation Methodology: Experience with frameworks like RAGAS or building custom “LLM-as-a-judge” systems.
  • check_circle Analytical Mindset: Ability to transform raw generation logs into actionable business metrics and technical insights.
  • check_circle Iterative Mindset: Focus on continuous product improvement through constant feedback loops.
  • check_circle Fluency in Russian and/or Ukrainian.

Preferred qualifications

  • Fine-Tuning: Practical experience in fine-tuning models for specific domain tasks or JSON compliance.
  • RAG Architecture: Understanding how to build and optimize Retrieval-Augmented Generation systems, including indexing, retrieval, and re-ranking.
  • Python: Basic knowledge for working with data science scripts or AI evaluation libraries.

Benefits

  • check_circle Remote Work Environment: Embrace the freedom to work from anywhere, anytime, promoting a healthy work-life balance.
  • check_circle Unlimited PTO: Enjoy unlimited paid time off to recharge and prioritize your well-being, without counting days.
  • check_circle Paid National Holidays: Celebrate and relax on national holidays with paid time off to unwind and recharge.
  • check_circle Company-provided MacBook: Experience seamless productivity with top-notch Apple MacBooks provided to all employees who need them.
  • check_circle Flexible Independent Contractor Agreement: Unlock the benefits of flexibility, autonomy, and entrepreneurial opportunities. Benefit from tax advantages, networking opportunities, reduced employment obligations, and the freedom to work from anywhere. Read more about it here: https://docs.google.com/document/d/1nkrN76JlZkbKj9WSOhlT1_mni_CZeDkHdwfIjPXVwvk/preview?tab=t.0#heading=h.ndsdl4wapxtt
  • check_circle Recruiter Screening (40 minutes)
  • check_circle Technical Interview (60 minutes)
  • check_circle Final Interview (30 minutes)

About the company

Ruby Labs is a leading tech company that creates and operates innovative consumer products. We offer a diverse range of opportunities across the health, education, and entertainment industries. Our innovative teams are driving the future of consumer-led products, and we're always looking for passionate individuals to join us. Learn more about our story at: https://rubylabs.com/about-us/

Tags & Focus Areas

Fulltime Remote Ai Ai Engineer Generative Ai

About Ruby Labs