Mark Liberman

Papers

2

Total Citations

15

H-Index

2

About

Mark Liberman is a foundational figure in computational linguistics and speech processing, best known as the founding director of the Linguistic Data Consortium (LDC) at the University of Pennsylvania. His primary research areas include corpus linguistics, phonetics, and the creation of large-scale digital language resources. Liberman’s most significant contribution is the establishment of the LDC in 1992, an open consortium that creates, distributes, and standardizes vast collections of speech and text databases, lexicons, and linguistic resources. This infrastructure has been critical for advancing automatic speech recognition, natural language processing, and empirical linguistics worldwide. His seminal 1998 paper on the creation and distribution of linguistic data (13 citations) outlines the LDC’s mission and impact, serving as a blueprint for open science in language research. Beyond this, Liberman has published influential work on prosody, discourse analysis, and computational phonology. He is also a prominent public intellectual, writing the widely-read blog “Language Log,” which brings linguistic insights to a broad audience. Through the LDC and his interdisciplinary scholarship, Liberman has shaped modern language technology and data-driven linguistics.

Research Focus

Key Achievements

2
H-Index
2
Papers
15
Total Citations
8
Avg Citations/Paper
🏆 Most Cited Paper
The creation, distribution and use of linguistic data: the case of the linguistic data consortium.
13 citations · 1998
📈 Most Prolific Year: 1998 (1 Papers)
🤝 Key Collaborators: 22

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 13 days ago