0% found this document useful (0 votes)

68 views3 pages

Ai DP 2

Natural language processing (NLP) is a field concerned with interactions between computers and human languages. NLP involves tasks like automatic summarization, machine translation, named entity recognition, and parsing. Modern NLP relies heavily on machine learning algorithms grounded in statistics. Major NLP tasks include understanding language at various levels like semantics, syntax, and discourse. Applications include question answering systems and conversational agents.

Uploaded by

Harsh Naraini

We take content rights seriously. If you suspect this is your content, claim it here.

Available Formats

Download as DOCX, PDF, TXT or read online on Scribd

0% found this document useful (0 votes)

68 views3 pages

Ai DP 2

Uploaded by

Harsh Naraini

We take content rights seriously. If you suspect this is your content, claim it here.

Available Formats

Download as DOCX, PDF, TXT or read online on Scribd

You are on page 1/ 3

Natural Language Processing

Natural language processing (NLP) is a field of computer science and linguistics concerned with the
interactions between computers and human (natural) languages. In theory, natural-language processing
is a very attractive method of human-computer interaction. Natural language understanding is
sometimes referred to as an AI-complete problem, because natural-language recognition seems to
require extensive knowledge about the outside world and the ability to manipulate it.

NLP has significant overlap with the field of computational linguistics, and is often considered a sub-field
of artificial intelligence.

Modern NLP algorithms are grounded in machine learning, especially statistical machine learning.
Research into modern statistical NLP algorithms requires an understanding of a number of disparate
fields, including linguistics, computer science, and statistics. For a discussion of the types of algorithms
currently used in NLP, see the article on pattern recognition.

Major tasks in NLP

The following is a list of some of the most commonly researched tasks in NLP. Note that some
of these tasks have direct real-world applications, while others more commonly serve as subtasks
that are used to aid in solving larger tasks. What distinguishes these tasks from other potential
and actual NLP tasks is not only the volume of research devoted to them but the fact that for
each one there is typically a well-defined problem setting, a standard metric for evaluating the
task, standard corpora on which the task can be evaluated, and competitions devoted to the
specific task.

 Automatic summarization: Produce a readable summary of a chunk of text. Often used to

provide summaries of text of a known type, such as articles in the financial section of a
newspaper.
 Coreference resolution: Given a sentence or larger chunk of text, determine which words
("mentions") refer to the same objects ("entities"). Anaphora resolution is a specific
example of this task, and is specifically concerned with matching up pronouns with the
nouns or names that they refer to. The more general task of coreference resolution also
includes identify so-called "bridging relationships" involving referring expressions. For
example, in a sentence such as "He entered John's house through the front door", "the
front door" is a referring expression and the bridging relationship to be identified is the
fact that the door being referred to is the front door of John's house (rather than of some
other structure that might also be referred to).
 Discourse analysis: This rubric includes a number of related tasks. One task is identifying
the discourse structure of connected text, i.e. the nature of the discourse relationships
between sentences (e.g. elaboration, explanation, contrast). Another possible task is
recognizing and classifying the speech acts in a chunk of text (e.g. yes-no question,
content question, statement, assertion, etc.).
 Machine translation: Automatically translate text from one human language to another.
This is one of the most difficult problems, and is a member of a class of problems
colloquially termed "AI-complete", i.e. requiring all of the different types of knowledge
that humans possess (grammar, semantics, facts about the real world, etc.) in order to
solve properly.
 Morphological segmentation: Separate words into individual morphemes and identify the
class of the morphemes. The difficulty of this task depends greatly on the complexity of
the morphology (i.e. the structure of words) of the language being considered. English
has fairly simple morphology, especially inflectional morphology, and thus it is often
possible to ignore this task entirely and simply model all possible forms of a word (e.g.
"open, opens, opened, opening") as separate words. In languages such as Turkish,
however, such an approach is not possible, as each dictionary entry has thousands of
possible word forms.
 Named entity recognition (NER): Given a stream of text, determine which items in the
text map to proper names, such as people or places, and what the type of each such name
is (e.g. person, location, organization). Note that, although capitalization can aid in
recognizing named entities in languages such as English, this information cannot aid in
determining the type of named entity, and in any case is often inaccurate or insufficient.
For example, the first word of a sentence is also capitalized, and named entities often
span several words, only some of which are capitalized. Furthermore, many other
languages in non-Western scripts (e.g. Chinese or Arabic) do not have any capitalization
at all, and even languages with capitalization may not consistently use it to distinguish
names. For example, German capitalizes all nouns, regardless of whether they refer to
names, and French and Spanish do not capitalize names that serve as adjectives.
 Natural language generation: Convert information from computer databases into readable
human language.
 Natural language understanding: Convert chunks of text into more formal representations
such as first-order logic structures that are easier for computer programs to manipulate.
 Optical character recognition (OCR): Given an image representing printed text,
determine the corresponding text.
 Part-of-speech tagging: Given a sentence, determine the part of speech for each word.
Many words, especially common ones, can serve as multiple parts of speech. For
example, "book" can be a noun ("the book on the table") or verb ("to book a flight");
"set" can be a noun, verb or adjective; and "out" can be any of at least five different parts
of speech. Note that some languages have more such ambiguity than others. Languages
with little inflectional morphology, such as English are particularly prone to such
ambiguity. Chinese is prone to such ambiguity because it is a tonal language during
verbalization. Such inflection is not readily conveyed via the entities employed within the
orthography to convey intended meaning.
 Parsing: Determine the parse tree (grammatical analysis) of a given sentence. The
grammar for natural languages is ambiguous and typical sentences have multiple possible
analyses. In fact, perhaps surprisingly, for a typical sentence there may be thousands of
potential parses (most of which will seem completely nonsensical to a human).
 Question answering: Given a human-language question, determine its answer. Typical
questions have a specific right answer (such as "What is the capital of Canada?"), but
sometimes open-ended questions are also considered (such as "What is the meaning of
life?").
 Relationship extraction: Given a chunk of text, identify the relationships among named
entities (i.e. who is the wife of whom).
 Sentence breaking (also known as sentence boundary disambiguation): Given a chunk of
text, find the sentence boundaries. Sentence boundaries are often marked by periods or
other punctuation marks, but these same characters can serve other purposes (e.g.
marking abbreviations).
 Speech recognition: Given a sound clip of a person or people speaking, determine the
textual representation of the speech. This is the opposite of text to speech and is one of
the extremely difficult problems colloquially termed "AI-complete" (see above). In
natural speech there are hardly any pauses between successive words, and thus speech
segmentation is a necessary subtask of speech recognition (see below). Note also that in
most spoken languages, the sounds representing successive letters blend into each other
in a process termed coarticulation, so the conversion of the analog signal to discrete
characters can be a very difficult process.
 Speech segmentation: Given a sound clip of a person or people speaking, separate it into
words. A subtask of speech recognition and typically grouped with it.
 Topic segmentation and recognition: Given a chunk of text, separate it into segments
each of which is devoted to a topic, and identify the topic of the segment.
 Word segmentation: Separate a chunk of continuous text into separate words. For a
language like English, this is fairly trivial, since words are usually separated by spaces.
However, some written languages like Chinese, Japanese and Thai do not mark word
boundaries in such a fashion, and in those languages text segmentation is a significant
task requiring knowledge of the vocabulary and morphology of words in the language.
 Word sense disambiguation: Many words have more than one meaning; we have to select
the meaning which makes the most sense in context. For this problem, we are typically
given a list of words and associated word senses, e.g. from a dictionary or from an online
resource such as WordNet.

In some cases, sets of related tasks are grouped into subfields of NLP that are often considered
separately from NLP as a whole. Examples include:

 Information retrieval (IR): This is concerned with storing, searching and retrieving
information. It is a separate field within computer science (closer to databases), but IR
relies on some NLP methods (for example, stemming). Some current research and
applications seek to bridge the gap between IR and NLP.
 Information extraction (IE): This is concerned in general with the extraction of semantic
information from text. This covers tasks such as named entity recognition, coreference
resolution, relationship extraction, etc.
 Speech processing: This covers speech recognition, text-to-speech and related tasks.

Tasks in NLP
No ratings yet
Tasks in NLP
7 pages
NLP for Information Retrieval
No ratings yet
NLP for Information Retrieval
8 pages
NLP Basics
No ratings yet
NLP Basics
7 pages
Notes
No ratings yet
Notes
9 pages
Project Report
No ratings yet
Project Report
12 pages
NLP Basics for Beginners
No ratings yet
NLP Basics for Beginners
4 pages
Session 14 - Computaional Linguistics
No ratings yet
Session 14 - Computaional Linguistics
23 pages
Khurana, D. (2017) - Natural Language Processing: State of Art, Current Trends and Challenges.
No ratings yet
Khurana, D. (2017) - Natural Language Processing: State of Art, Current Trends and Challenges.
25 pages
Selected Topic CH 1
No ratings yet
Selected Topic CH 1
36 pages
7
No ratings yet
7
4 pages
Introduction To NLP
No ratings yet
Introduction To NLP
37 pages
NLP Chapter-1
No ratings yet
NLP Chapter-1
24 pages
NLP Module1-4
No ratings yet
NLP Module1-4
100 pages
NLP Assignment 1
No ratings yet
NLP Assignment 1
4 pages
NLP Trends and Challenges
No ratings yet
NLP Trends and Challenges
23 pages
NLP Ambiguity
No ratings yet
NLP Ambiguity
35 pages
Unit1 (Part1)
No ratings yet
Unit1 (Part1)
49 pages
Unit V Expert Systems Notes
No ratings yet
Unit V Expert Systems Notes
15 pages
CH 1 Introduction To NLP
No ratings yet
CH 1 Introduction To NLP
31 pages
NLP Unit 1,2 Notes
No ratings yet
NLP Unit 1,2 Notes
37 pages
NLP for AI and Computer Science Enthusiasts
No ratings yet
NLP for AI and Computer Science Enthusiasts
27 pages
Assignment of AI Finished
No ratings yet
Assignment of AI Finished
16 pages
NLP: Techniques and Applications
No ratings yet
NLP: Techniques and Applications
45 pages
Unit V Intelligence and Applications: Morphological Analysis/Lexical Analysis
No ratings yet
Unit V Intelligence and Applications: Morphological Analysis/Lexical Analysis
30 pages
Natural Language Processing
No ratings yet
Natural Language Processing
17 pages
Linguistic Issues and Methods in Computa A4919075
No ratings yet
Linguistic Issues and Methods in Computa A4919075
7 pages
Adaca 2012
No ratings yet
Adaca 2012
97 pages
Natural Language Processing
100% (2)
Natural Language Processing
9 pages
Lecture 1: Introduction To NLP: Understand Concepts Applications
No ratings yet
Lecture 1: Introduction To NLP: Understand Concepts Applications
32 pages
0 Unit-1 Introducntion To NLP
No ratings yet
0 Unit-1 Introducntion To NLP
41 pages
Computational Linguistics in The Netherlands 2000 Jorn Veenstra Download
No ratings yet
Computational Linguistics in The Netherlands 2000 Jorn Veenstra Download
147 pages
Unit 1-NLP
No ratings yet
Unit 1-NLP
62 pages
3nlp Computer
No ratings yet
3nlp Computer
83 pages
Feature Eng
No ratings yet
Feature Eng
34 pages
NLP Trends and Challenges
No ratings yet
NLP Trends and Challenges
26 pages
An In-Depth Exploration of Natural Language Processing: Evolution, Applications, and Future Directions
100% (8)
An In-Depth Exploration of Natural Language Processing: Evolution, Applications, and Future Directions
5 pages
Corpus Linguistics 1
No ratings yet
Corpus Linguistics 1
48 pages
NLP Introduction Overview
No ratings yet
NLP Introduction Overview
34 pages
Chapter 6
100% (1)
Chapter 6
28 pages
Natural Language Processing
No ratings yet
Natural Language Processing
32 pages
NLP: Trends, Challenges & Insights
No ratings yet
NLP: Trends, Challenges & Insights
32 pages
Unit 4
No ratings yet
Unit 4
16 pages
(A) What Is Traditional Model of NLP?: Unit - 1
No ratings yet
(A) What Is Traditional Model of NLP?: Unit - 1
18 pages
NLP and ES
No ratings yet
NLP and ES
1 page
Definition: Natural Language Processing Is A Theoretically Motivated Range of Computational
No ratings yet
Definition: Natural Language Processing Is A Theoretically Motivated Range of Computational
14 pages
1.introduction To Natural Language Processing (NLP)
100% (1)
1.introduction To Natural Language Processing (NLP)
37 pages
Chapter 5 - Communication Perceving and Acting
No ratings yet
Chapter 5 - Communication Perceving and Acting
20 pages
Natural Language Processing State of The Art Curre
No ratings yet
Natural Language Processing State of The Art Curre
26 pages
NLP Unit 1 Answers
No ratings yet
NLP Unit 1 Answers
7 pages
NLP Notes
No ratings yet
NLP Notes
16 pages
Natural Language Processing State of The Art Curre
No ratings yet
Natural Language Processing State of The Art Curre
26 pages
NLP Unit I Notes-1
75% (4)
NLP Unit I Notes-1
22 pages
NLP Applications in Healthcare
No ratings yet
NLP Applications in Healthcare
71 pages
NLP One Mark Questions With Answers
No ratings yet
NLP One Mark Questions With Answers
8 pages
NLP Techniques: POS & Semantic Tagging
No ratings yet
NLP Techniques: POS & Semantic Tagging
30 pages
Natural Language Processing Tools and Approaches
No ratings yet
Natural Language Processing Tools and Approaches
106 pages
Lecture 8 - Pre Processing Techniques
No ratings yet
Lecture 8 - Pre Processing Techniques
14 pages
Belonging3477formative Final
No ratings yet
Belonging3477formative Final
43 pages
Gateway B1plus Students Book Answers - B1+ 1 of 18 4 Students' Own Answers 5 Hold His - Studocu
No ratings yet
Gateway B1plus Students Book Answers - B1+ 1 of 18 4 Students' Own Answers 5 Hold His - Studocu
1 page
Alphanumeric Test For Cuet
No ratings yet
Alphanumeric Test For Cuet
6 pages
Grade - 7 (Text Book)
No ratings yet
Grade - 7 (Text Book)
9 pages
Alike English Vi STD Interview Guide
No ratings yet
Alike English Vi STD Interview Guide
83 pages
12 Tenses in English Grammar Verb Tenses PDF
75% (8)
12 Tenses in English Grammar Verb Tenses PDF
6 pages
Unit 5 Animals - Lesson 3.1
No ratings yet
Unit 5 Animals - Lesson 3.1
40 pages
English Accents & Dialects Guide
No ratings yet
English Accents & Dialects Guide
6 pages
Class Worksheet (Sentence Making Practice)
No ratings yet
Class Worksheet (Sentence Making Practice)
10 pages
Direct & Indirect Speech
No ratings yet
Direct & Indirect Speech
22 pages
Ahmed and Mohammadzadeh-Speech and Thought Presentation in Munro (2023)
100% (1)
Ahmed and Mohammadzadeh-Speech and Thought Presentation in Munro (2023)
9 pages
Modal Verbs Learning Guide
25% (4)
Modal Verbs Learning Guide
25 pages
Lista Verbe Neregulate Engleză
100% (8)
Lista Verbe Neregulate Engleză
61 pages
Positive Sentences: Who? Form of Verb Examples
No ratings yet
Positive Sentences: Who? Form of Verb Examples
3 pages
Cambridge O Level: English Language 1123/11 October/November 2022
No ratings yet
Cambridge O Level: English Language 1123/11 October/November 2022
12 pages
Living Language 2nd Edition Laura M. Ahearn PDF Download
No ratings yet
Living Language 2nd Edition Laura M. Ahearn PDF Download
53 pages
3.19 Passive Voice
No ratings yet
3.19 Passive Voice
10 pages
Language and Culture
100% (1)
Language and Culture
75 pages
EEF SM 1 Unit 3
No ratings yet
EEF SM 1 Unit 3
3 pages
Grammar Activity 2A-2B
100% (1)
Grammar Activity 2A-2B
2 pages
CORE Class Grade 1 (Phonics Survey)
No ratings yet
CORE Class Grade 1 (Phonics Survey)
1 page
Seloka: Jurnal Pendidikan Bahasa Dan Sastra Indonesia: Susilo Rini & Wagiran
No ratings yet
Seloka: Jurnal Pendidikan Bahasa Dan Sastra Indonesia: Susilo Rini & Wagiran
7 pages
Test It Fix It English Grammar Pre-Intermedia
No ratings yet
Test It Fix It English Grammar Pre-Intermedia
45 pages
Language Functions Explained
No ratings yet
Language Functions Explained
4 pages
Prefixes & Suffixes: Meanings & Examples
No ratings yet
Prefixes & Suffixes: Meanings & Examples
31 pages
Delta Module One Reading Guide
100% (1)
Delta Module One Reading Guide
3 pages
Thesis About The Importance of English Language
100% (3)
Thesis About The Importance of English Language
8 pages
Past Simple Test: A. Write The Past Forms of The Regular and Irregular Verbs
No ratings yet
Past Simple Test: A. Write The Past Forms of The Regular and Irregular Verbs
4 pages
Figurative Language Creates Figures
No ratings yet
Figurative Language Creates Figures
2 pages
Communication Skills New Syllabus
No ratings yet
Communication Skills New Syllabus
5 pages

Ai DP 2

Uploaded by

Ai DP 2

Uploaded by

Natural Language Processing

Major tasks in NLP

 Automatic summarization: Produce a readable summary of a chunk of text. Often used to

You might also like