Logo for Pinecone

Senior/Staff Software Engineer, Search & Retrieval Infrastructure

Role overview

Qualifications

  • Proven track record (typically 6+ years) of shipping production-grade backends for large-scale systems
  • Experience in semantic search, vector databases, hybrid retrieval strategies
  • Expert in at least one major programming language like Go, Rust, C++, Java, or Python
  • Familiarity with modern infrastructure tools like Kubernetes, Terraform, or Pulumi

Responsibilities

  • Design and build scalable platform components leveraging advanced retrieval via query planning, semantic and hybrid search
  • Build backend services for semantic and hybrid retrieval, knowledge graph construction, and retrieval orchestration
  • Improve retrieval quality through evaluation and observability frameworks
  • Design APIs for internal and external user and agentic consumers

About the company

Pinecone logo

Pinecone

Cloud Computing & Infrastructure (IaaS/PaaS)

Pinecone is a fully managed vector database that makes it easy to add vector search to production applications. It combines state-of-the-art vector search libraries, advanced features such as filtering, and distributed infrastructure to provide high performance and reliability at any scale. No more hassles of benchmarking and tuning algorithms or building and maintaining infrastructure for vector search. We are engineers who built large machine learning platforms, databases, and search engines at AWS, Yahoo, Splunk, Databricks, and more. We are scientists who transformed businesses with ML applications such as shopping recommendations, online advertising, semantic search, anomaly detection, and more. We are researchers who collectively published more than 100 academic papers and patents on machine learning, systems, and algorithms.

Company details

Company typeScaleup
IndustryCloud Computing & Infrastructure (IaaS/PaaS)
Company size51 - 200

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

About Pinecone

Pinecone is the knowledge infrastructure for AI at scale. Its leading vector database and knowledge engine, Pinecone Nexus, power accurate, performant AI applications for more than 9,000 customers and 800,000 developers worldwide. Pinecone's mission is to make AI knowledgeable. Pinecone is based in New York and raised $138M in funding from Andreessen Horowitz, ICONIQ, Menlo Ventures, and Wing Venture Capital.

About the Team and Role:

We are hiring a senior/staff software engineer to help design and build core components of our next-generation knowledge retrieval system built for the AI era – search and retrieval infrastructure that powers high-quality, scalable, and enterprise-grade agentic systems. You’ll build the framework that allows our customers to connect knowledge–synthesized from structured and unstructured data–to modern LLM-powered applications, leveraging the world’s best-in-class vector DB supporting semantic search and hybrid retrieval. This role is ideal for someone who loves backend system architecture, distributed systems, and applied AI infrastructure. It is a high impact role with significant ownership across architecture, performance, and system reliability.

Responsibilities:

  • Design and build scalable platform components leveraging advanced retrieval via query planning, semantic and hybrid search, metadata-aware search, and LLM generation

  • Design and build optimized indexing pipelines for structured and unstructured data

  • Build backend services for semantic and hybrid retrieval, knowledge graph construction, and retrieval orchestration

  • Improve retrieval quality through evaluation and observability frameworks

  • Design APIs for internal and external user and agentic consumers

  • Optimize latency, throughput and cost across large-scale inference and retrieval workloads

  • Drive technical direction for reliability and security

What You’ll Bring to the Table:

To thrive in this role, you don't need to check every single box, but you should be deeply passionate about how to turn data into knowledge.

Systems Expertise

  • Architectural Depth: You have a proven track record (typically 6+ years) of shipping production-grade backends for large-scale systems. You don’t just write code; you design for high throughput, low latency, and long-term maintainability.

  • Data Engineering Savvy: You’re comfortable building high-throughput indexing pipelines that handle both the messy world of unstructured data and the rigid world of structured schemas.

AI & Retrieval

  • Retrieval Intuition: You understand that "search" is more than just a keyword match. You have direct experience (or deep theoretical knowledge) in semantic search, vector databases, hybrid retrieval strategies, or with traditional search engines like Elastic or OpenSearch.

  • RAG & Orchestration: You understand the nuances of Retrieval-Augmented Generation (RAG) patterns, from embedding pipelines and hybrid search techniques to how query planning and metadata filtering can make or break an LLM's performance.

Technical

  • Language Fluency: You are an expert in at least one major language like Go, Rust, C++, Java, or Python.

  • Infrastructure: Familiarity and experience with modern infrastructure tools, such as Kubernetes, cloud-native architectures, and observability frameworks, as well as infrastructure-as-code tools like Terraform or Pulumi.

Ownership & Impact

  • Product Thinking: You don't just build to spec; you build for the user. You can design clean, intuitive APIs that both human developers and autonomous agents will love.

  • Ambiguity Navigator: You’re comfortable in a high-growth environment. You prefer "owning a problem" over "executing a ticket."

Bonus Points

  • Experience building multi-tenant SaaS platforms.

  • Experience with retrieval evaluation frameworks—knowing how to actually measure "good" search results.

  • Experience with query planning or agentic reasoning loops (e.g., teaching a system how to break down a complex prompt into multiple specific steps).

Perks & Benefits:

  • Comprehensive health coverage including medical, dental, vision, and mental health resources

  • 401(k) Plan

  • Equity award

  • Flexible time off

  • Paid parental leave

  • Annual Company Retreat

  • WFH Equipment Stipend

All qualified applicants will receive considerations for employment without regard to race, color, religion, sex, age, disability, marital status, familial status, sexual orientation, pregnancy, gender identity, gender expression, national origin, ancestry, citizenship status, veteran status, and any other legally protected status under federal, state, or local anti-discrimination laws.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Software Engineer Related jobs

Other jobs at Pinecone

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.