LLMpediaThe first transparent, open encyclopedia generated by LLMs

Andrei Broder

Note: This article was automatically generated by a large language model (LLM) from purely parametric knowledge (no retrieval). It may contain inaccuracies or hallucinations. This encyclopedia is part of a research project currently under review.
Article Genealogy
Parent: Hamming distance Hop 5 terminal

This article was accepted into the corpus but its outbound wikilinks were never NER-processed — typical at the deepest BFS hop or when the run's entity cap was reached. No expansion funnel to show.

Andrei Broder
NameAndrei Broder
FieldsComputer science, Information retrieval, Algorithms
WorkplacesYahoo!, Google, AltaVista, IBM
Alma materPolytechnic University of Bucharest, Stony Brook University
Known forShingling, min-hash, web crawling, search advertising

Andrei Broder

Andrei Broder is a computer scientist known for foundational work in information retrieval, web search, and algorithms that shaped large-scale Internet services. He contributed to techniques such as shingling and min-hash that influenced systems developed at organizations including AltaVista, IBM Research, Yahoo!, and Google. Broder's research intersected with practitioners and theorists across institutions such as Stony Brook University, Massachusetts Institute of Technology, Stanford University, and companies like Microsoft and Amazon.com.

Early life and education

Broder studied at the Polytechnic University of Bucharest in Romania before pursuing graduate studies in the United States. He completed advanced degrees at Stony Brook University, where he worked on topics linking theory and practice in algorithms and data structures. During his doctoral and postdoctoral period he interacted with researchers from institutions including Bell Labs, IBM Research, AT&T, and University of California, Berkeley, fostering collaborations that later influenced projects at Compaq and Hewlett-Packard.

Research and contributions

Broder's research produced techniques widely cited across literature in information retrieval, data mining, and computer science theory. He co-developed shingling and min-hash methods for near-duplicate detection, which were adopted in systems at Google's web crawler efforts, Yahoo!'s search infrastructure, and by groups at Microsoft Research. His work on locality-sensitive hashing linked to research from Piotr Indyk, Rajeev Motwani, and others in the context of approximate nearest neighbors used in projects at Stanford University and MIT.

Broder introduced taxonomy and classification of web queries—often described as the "broder taxonomy"—that categorized queries into navigational, informational, and transactional classes, influencing teams at Google and Microsoft Bing and informing research at Carnegie Mellon University and Yahoo! Research. His analyses of web graph structure connected to seminal studies by Jon Kleinberg, Sergey Brin, and Larry Page on link analysis and the PageRank algorithm; those connections informed work on web crawling, duplicate detection, and spam mitigation used by AltaVista and later search engines.

His publications explored search advertising, click-through modeling, and sponsored search mechanics, intersecting with work at Overture Services and advertising groups at Google AdWords and Microsoft Advertising. Broder's collaborative studies on large-scale indexing, compression, and distributed computing drew upon methods used in MapReduce deployments at Google and distributed storage designs at Amazon Web Services and Yahoo!.

Career and positions

Broder held research and leadership roles spanning academia and industry. He worked at AltaVista during its era as a pioneering search engine and later held positions at IBM Research, contributing to projects that combined theoretical foundations with industrial-scale systems. He joined Yahoo! in senior research and engineering capacities, influencing search, advertising, and infrastructure. Broder later moved to Google, where he worked on core search technologies and advertising systems alongside teams that included researchers from Stanford University, MIT, and UC Berkeley.

Throughout his career Broder collaborated with scholars and practitioners at institutions such as Columbia University, Princeton University, University of Massachusetts Amherst, and corporate research labs including Bell Labs and Microsoft Research. He participated in program committees for conferences like SIGIR, KDD, WWW, and SODA, and contributed to workshops hosted by ACM and IEEE.

Awards and honors

Broder received recognition from professional societies and peers for his impact on information retrieval and web search engineering. His papers have been highly cited in venues including ACM SIGIR, WWW Conference, and KDD Conference, and he has been invited to serve as an industry fellow and advisor to initiatives at National Science Foundation-funded centers and collaborations with universities like Harvard University and Yale University. He has been acknowledged by organizations such as ACM and IEEE for his contributions to scalable algorithms and search technologies.

Selected publications

- Broder, A., et al., papers on shingling and min-hash techniques, influential in information retrieval and data mining communities; cited in works from Stanford University and MIT researchers. - Broder, A., publications on web query classification (navigational, informational, transactional) used by teams at Google, Microsoft, and Yahoo! Research. - Broder, A., studies of web graph structure and duplicate detection connected to research by Jon Kleinberg, Sergey Brin, and Larry Page. - Broder, A., collaborative work on search advertising, click-through modeling, and sponsored search, referenced by groups at Google AdWords and Microsoft Advertising. - Broder, A., contributions to large-scale indexing, compression, and distributed computing, informing engineering at Amazon Web Services, Yahoo!, and Google.

Category:Computer scientists Category:Information retrieval researchers