← Back to Jobs
Databricks logo
1 week agogreenhouse

Staff Software Engineer - Search Platform

Databricks / Software Development
Bengaluru, Karnataka, IndiaOn-siteFull Time
Apply on company site
AI Summary

Lead the development and deployment of ML-based search and discovery relevance models and systems. Collaborate with cross-functional teams to design automated pipelines for query understanding, ranking, and model evaluation.

Requires a BS degree or higher in Computer Science and 10+ years of experience in search relevance systems at scale. Candidates must have expertise in applying LLMs to search and strong proficiency in NLP or query understanding.

Job Details
Job Type: FULL TIME
Visa Sponsorship: No
Education: bachelor degree, postgraduate degree
Work Arrangement: On-site
Experience Level: 10+ years
Hours: 40 hrs/week
Language: English
Key Skills
Search relevanceMachine learningNLPQuery understandingRankingRetrievalLLMData preprocessingModel evaluationPythonSoftware engineeringDistributed systemsText miningRecommendation systemsConversational AI
Insider connections @Databricks
Members only

Find people at Databricks who may share hiring insights or referrals for this role.

Meet the people behind the opportunity

Create a free account to search for hiring managers and potential referral contacts.

No contact data is shown on this public page.

P-375

The Search team at Databricks sits at the forefront of advancing AI/ML-powered products. Databricks’ customers are continuously creating new assets (tables, notebooks, dashboards, datarooms, pipelines, sql queries, ml models etc.) on the platform. Finding an asset is a critical user journey for Databricks’ customers which helps them accomplish their tasks. 

As our Search product continues to evolve, we are seeking a Staff Engineer to lead enhancements to our Search Quality. We are focusing on enhancing search ranking, improving query understanding, building robust evals and growing the coverage of assets to enable seamless search at scale.

Key Responsibilities

  • Drive the development and deployment of ML based search and discovery relevance models and systems integrated with Databricks' products and services. 
  • Design and implement automated ML and NLP pipelines for data preprocessing, query understanding and rewrite, ranking and retrieval, and model evaluation, enabling rapid experimentation and iteration. 
  • Collaborate with product managers and cross-functional teams to drive technology-first initiatives that enable novel business strategies and product roadmaps for the search and discovery experience. 
  • Contribute to building a robust framework for evaluating search ranking improvements - both offline and online.

 

What We’re Looking For

  • BS+ (M.S. or PhD preferred) in Computer Science, or a related field.
  • 10+ years experience developing search relevance systems at scale in production or in high-impact research environments.
  • Experience applying LLM to search relevance
  • Experience in one or more of the following:  
    • Query understanding
    • NLP
    • Text mining
    • Recommendations
    • Personalization
    • Discovery
    • Conversational AI 
  • Strong understanding of computer science fundamentals.
  • Contributions to well-used open-source projects. 

 

Why Join Us?
At Databricks, we are building state-of-the-art AI solutions that redefine how users interact with data and our products. You’ll have the opportunity to shape the future of AI-driven products at Databricks, work with cutting-edge models, and collaborate with a world-class team of AI and ML experts.

If you're excited about pushing the boundaries of AI in real-world applications, we’d love to hear from you!

About Databricks

Databricks is the Data and AI company. More than 20,000 organizations worldwide — including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 — rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram.

Benefits

At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here.

Our Commitment to Diversity and Inclusion

At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics.

Compliance

If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.

About Databricks

Databricks empowers over 10,000 global organizations to harness their data with AI through its innovative Data Intelligence Platform, founded by the creators of Apache Spark™.

Software Development
16,344 employees
San Francisco, CA
Most staff in USA
$32B raised
Employee ratings
From Glassdoor · 1.8k reviews · View profile
Compensation & benefits4.2
Career opportunities4.0
Diversity & inclusion3.9
Culture & values3.8
Senior management3.7
Work–life balance3.4
76%
would recommend to a friend
86%
positive business outlook
Categories
Data StorageInformation TechnologyData IntegrationArtificial IntelligenceSoftwareData Management
Specialties
Apache SparkApache Spark TrainingCloud ComputingBig DataData ScienceDelta LakeData LakehouseMLflowMachine LearningData EngineeringData WarehousingData StreamingOpen SourceGenerative AIArtificial IntelligenceData IntelligenceData ManagementData GoveranceGenerative AIand AI/ML Ops