Job: Live Search Core Relevance group, Microsoft

Einat Amitay <einat-7z/[email protected]> Tue, 7 Aug 2007 13:07:40 +0300
Newsgroups gmane.comp.information-retrieval.webir
Message-ID <OFE8D70ACB.C92E8119-ONC2257330.003775BA-C2257330.0037A276@il.ibm.com>
Redmond, WA: Software Engineer in Live Search Data Mining team at
Microsoft, Live Search Core Relevance group

The Live Search Core Relevance group is hiring extremely talented, highly
motivated and productive individuals. You will be part of the Live Search
Data Mining team, developing advanced and practical data mining and machine
learning techniques for solving the hottest and most challenging problems
in the world. Here, you have the right environment and strong support to
drive your favorite features to solution. You are empowered to influence
millions of end users while impacting Windows Live’s company value. You
will have the opportunities to work together with world class researchers
and developers to stay in the front of advancing technology.

The goal of Live Search is to deliver the most relevant internet search to
our customers by building the best search engine. To achieve this goal,
data mining/ machine learning/ statistics/ database techniques are
critical. We have a huge amount of data ranging from user interaction logs,
to web documents, to user feedback, to system performance data. This data
contains valuable information for building the #1 system to serve people’s
information needs. The main challenge is to effectively discover this
information and use it in the system. We are developing advanced and
practical data mining and machine learning techniques to derive information
(such as: understanding user intent, discovering patterns to improve search
relevance and result ranking, detecting and filtering bot traffic, spam,
and noise, measuring relevance and user satisfaction quickly and
effectively, identifying new information needs from end users, identifying
weak points of the search system, discovering system performance problems,
and many many more). We also work on techniques to manage and process
petabytes of data efficiently and effectively. You will have opportunities
to use and contribute to a high performance super-parallel/distributed
massive storage computer system.

Job Responsibilities include:

Analyze (including processing) a huge amount of data by using data mining
and statistics techniques. The goal is to generate actionable insights for
improving search technology and systems, increasing user satisfaction, as
well as understanding search user behavior and Search business.
Research and exploration in the areas of data mining, machine learning,
text mining, search system live measurement and monitoring, bot/spam
traffic identification and filtering, etc.
Work with other teams in Live Search on identifying problems in different
areas where data mining/machine learning/statistics can help. Explore and
develop solutions to these problems. Act as an expert in the area of data
mining/machine learning/statistics to serve the fast growing needs of Live
Search.
Develop techniques/algorithms for research and analysis work mentioned
above.
Design and carry out experiments to evaluate research results and their
real impact on Live Search production systems.
Design and develop software systems/solutions to push research and analysis
results into production systems and generate impacts on very large number
of users.
Own some of the features/problem spaces in this area and provide technical
leadership to other developers.

Qualifications:

Extensive knowledge and experience in data mining/text mining/machine
learning/statistics/databases
Strong theory/algorithm background and very good understanding on how to
apply advanced knowledge to solve real problems
Solid experience in very large real world data analysis especially web data
analysis
Experience/knowledge with various data analysis tools, data mining tools,
and statistical packages
Superior communications skills, both verbal and written
Ability to work independently and in a team to research innovative
solutions to challenging business/technical problems
Attention to detail and data accuracy
Extensive software design and development skills/experience with
C/C++/C#/Perl/SQL
Minimum of 7 years in the industry.
Master degree or PhD (preferred) in the area of data mining/machine
learning/statistics is required.

Contact:
Heather McGough, [email protected]