UNIT-2 notes 1

You might also like

Download as docx, pdf, or txt
Download as docx, pdf, or txt
You are on page 1of 12

CS6503 – Theory of Computation

UNIT-2
GRAMMARS
Define CFG.

A context-free grammar (CFG) consisting of a finite set of grammar rules is a


quadruple (N, T, P, S) where

 N is a set of non-terminal symbols.


 T is a set of terminals where N ∩ T = NULL.
 P is a set of rules
 S is the start symbol

Define production rule.

A set of production rules that describe all possible strings in a given formal language.
Production rules are simple replacements. For example, the rule

P = S → SS | aSb | ε

Define terminal and non terminal symbols.

Terminals
A terminal is a symbol which does not appear on the left-hand side of any production. A
grammar contains a set of terminal symbols (tokens) such as +, , *, and other tokens
Nonterminals
Nonterminals are the non-leaf nodes in a parse tree. In the Expression grammar, E, T, and F
are nonteminals.

Write about the types of grammars.

What is ambiguity?
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
If a context free grammar G has more than one derivation tree for some string w ∈ L(G), it
is called an ambiguous grammar. There exist multiple right-most or left-most derivations for some
string generated from that grammar.

Define sentential form.

A sentential form is any string derivable from the start symbol.

Define parse tree.

A parse tree is an ordered, rooted tree that represents the syntactic structure of a string
according to some context-free grammar.

What is a derivation?

A derivation of a string for a grammar is a sequence of grammar rule applications that


transforms the start symbol into the string. A derivation proves that the string belongs to the
grammar's language.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
What is null production and unit production?

Unit Productions
Any production rule in the form A → B where A, B ∈ Non-terminal is called unit
production.
Null Productions
In a CFG, a non-terminal symbol ‘A’ is a nullable variable if there is a production A → ϵ or
there is a derivation that starts at A and finally ends up with ϵ: A → .......… → ϵ

What are the two normal forms of CFG?


Chomsky Normal Form
A CFG is in Chomsky Normal Form if the Productions are in the following forms:
A→a
 A → BC
S→ϵ
where A, B, and C are non-terminals and a is terminal.
Greibach Normal Form
A CFG is in Greibach Normal Form if the Productions are in the following forms:
A→b
A → bD1…Dn
S→ϵ
where A, D1,....,Dn are non-terminals and b is a terminal.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
Derivations and Languages

A derivation tree or parse tree is an ordered rooted tree that graphically represents the
semantic information a string derived from a context-free grammar.
Representation Technique:
1. Root vertex: Must be labeled by the start symbol.
2. Vertex: Labeled by a non-terminal symbol.
3. Leaves: Labeled by a terminal symbol or ε.

There are two different approaches to draw a derivation tree:


1. Top-down Approach:
(a) Starts with the starting symbol S
(b) Goes down to tree leaves using productions
2. Bottom-up Approach:
(a) Starts from tree leaves
(b) Proceeds upward to the root which is the starting symbol S
Representation of Derivations:
 Sentential Form
 Parser Tree
(i)sentential form.

A sentential form is any string derivable from the start symbol.


CS6503 – Theory of Computation
UNIT-2
GRAMMARS
Parse tree.

A parse tree is an ordered, rooted tree that represents the syntactic structure of a string
according to some context-free grammar.

Leftmost and Rightmost Derivation of a String


 Leftmost derivation - A leftmost derivation is obtained by applying production to the
leftmost variable in each step.
 Rightmost derivation - A rightmost derivation is obtained by applying production to the
rightmost variable in each step.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
Ambiguous Grammar

If a context free grammar G has more than one derivation tree for some string w ∈ L(G),
it is called an ambiguous grammar. There exist multiple right-most or left-most derivations for
some string generated from that grammar.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
CFG Simplification

Removal of Useless Symbols


Any symbol is useful when it appears on the right hand side, the production rule and
generate some terminal string. If no such derivation exists then it is supposed to be an useless
symbols.

Removal of Unit Productions

Any production rule in the form A → B where A, B ∈ Non-terminal is called unit production.
Removal Procedure −
Step 1 − To remove A → B, add production A → x to the grammar rule whenever B → x
occurs in the grammar. [x ∈ Terminal, x can be Null]
Step 2 − Delete A → B from the grammar.
Step 3 − Repeat from step 1 until all unit productions are removed.

Removal of Null Productions

In a CFG, a non-terminal symbol ‘A’ is a nullable variable if there is a production A → ε or there is


a derivation that starts at A and finally ends up with
ε: A → .......… → ε
Removal Procedure
Step 1 − Find out nullable non-terminal variables which derive ε.
Step 2 − For each production A → a, construct all productions A → x where x is obtained
from ‘a’ by removing one or multiple non-terminals from Step 1.
Step 3 − Combine the original productions with the result of step 2 and remove ε -
productions.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
Chomsky Normal Form

A CFG is in Chomsky Normal Form if the Productions are in the following forms −
A→a
A → BC
S→ε
where A, B, and C are non-terminals and a is terminal.

Algorithm to Convert into Chomsky Normal Form –

Step 1 − If the start symbol S occurs on some right side, create a new start symbol S’ and a new
production S’→ S.
Step 2 − Remove Null productions.
Step 3 − Remove unit productions.
Step 4 − Replace each production A → B1…Bn where n > 2 with A → B1Cwhere C → B2 …Bn.
Repeat this step for all productions having two or more symbols in the right side.
Step 5 − If the right side of any production is in the form A → aB where a is a terminal and A,
B are non-terminal, then the production is replaced by A → XB and X → a. Repeat this step for
every production which is in the form A → aB.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
CS6503 – Theory of Computation
UNIT-2
GRAMMARS

Greibach Normal Form

A CFG is in Greibach Normal Form if the Productions are in the following forms −

A→b
A → bD1…Dn
S→ε
where A, D1,....,Dn are non-terminals and b is a terminal.
Algorithm to Convert a CFG into Greibach Normal Form
Step 1 − If the start symbol S occurs on some right side, create a new start symbol S’ and a new
production S’ → S.
Step 2 − Remove Null productions. (Using the Null production removal algorithm discussed
earlier)
Step 3 − Remove unit productions. (Using the Unit production removal algorithm discussed
earlier)
Step 4 − Remove all direct and indirect left-recursion.
Step 5 − Do proper substitutions of productions to convert it into the proper form of GNF.
CS6503 – Theory of Computation
UNIT-2
GRAMMARS
CS6503 – Theory of Computation
UNIT-2
GRAMMARS

You might also like