ParseKeggCompoundManager
extends ParseKeggAbstractManager
in package
Class ParseKeggCompoundManager A compound record describes one metabolite : its formula, the reactions it takes part in, the enzymes acting on it and the pathways it belongs to.
Tags
Table of Contents
Properties
- $entry : string
- $names : array<string|int, mixed>
- $dbLinks : array<string|int, mixed>
- $enzymes : array<string|int, mixed>
- EC numbers of the enzymes acting on the compound.
- $formula : string
- $pathways : array<string|int, mixed>
- $reactions : array<string|int, mixed>
Methods
- __construct() : mixed
- Constructor.
- getDbLinks() : array<string|int, mixed>
- getEntry() : string
- getEntryId() : string
- Extracts the identifier uniquely naming a KEGG record.
- getEnzymes() : array<string|int, mixed>
- getFormat() : string
- The name this format is known by in the collection records and in DatabaseParserFactory.
- getFormula() : string
- getNames() : array<string|int, mixed>
- getPathways() : array<string|int, mixed>
- getReactions() : array<string|int, mixed>
- isEntryEnd() : bool
- Tells whether a line closes a KEGG record. KEGG closes on three slashes where most flat files use two.
- isEntryStart() : bool
- Tells whether a line opens a new KEGG record.
- parseDataFile() : Sequence
- Parses a KEGG compound data file and populates this manager's fields.
- joinLines() : string
- Joins the lines of a field into one string. A line opening with "$" continues the word the line above broke off, so it joins without a space.
- parseDbLinks() : array<string|int, mixed>
- Reads a DBLINKS field into pairs of database name and identifier, one per line.
- parsePathways() : array<string|int, mixed>
- Reads a PATHWAY field into pairs of map identifier and pathway name.
- readData() : string
- Reads the data a line carries, which starts at its thirteenth column.
- readEntryId() : string
- Reads the identifier out of an ENTRY line. Its last word names the kind of record rather than the record itself - "C00031 Compound", "EC 2.7.1.1 Enzyme" - so the identifier is what comes before it, which for an enzyme is the two words "EC" and its number.
- readFields() : array<string|int, mixed>
- Gathers a record into its fields : one entry per label, holding the data of its own line and of every continuation line below it.
- readLabel() : string
- Reads the label a line carries in its first twelve columns.
- splitTokens() : array<string|int, mixed>
- Splits the lines of a field into the whitespace-separated identifiers they list.
Properties
$entry
protected
string
$entry
= ""
$names
protected
array<string|int, mixed>
$names
= []
$dbLinks
private
array<string|int, mixed>
$dbLinks
= []
$enzymes
EC numbers of the enzymes acting on the compound.
private
array<string|int, mixed>
$enzymes
= []
$formula
private
string
$formula
= ""
$pathways
private
array<string|int, mixed>
$pathways
= []
$reactions
private
array<string|int, mixed>
$reactions
= []
Methods
__construct()
Constructor.
public
__construct() : mixed
getDbLinks()
public
getDbLinks() : array<string|int, mixed>
Return values
array<string|int, mixed>getEntry()
public
getEntry() : string
Return values
stringgetEntryId()
Extracts the identifier uniquely naming a KEGG record.
public
static getEntryId(array<string|int, mixed> $aFlines, string $sLine) : string
Parameters
- $aFlines : array<string|int, mixed>
-
The whole file, buffered
- $sLine : string
-
The line opening the entry
Return values
stringgetEnzymes()
public
getEnzymes() : array<string|int, mixed>
Return values
array<string|int, mixed>getFormat()
The name this format is known by in the collection records and in DatabaseParserFactory.
public
static getFormat() : string
Return values
stringgetFormula()
public
getFormula() : string
Return values
stringgetNames()
public
getNames() : array<string|int, mixed>
Return values
array<string|int, mixed>getPathways()
public
getPathways() : array<string|int, mixed>
Return values
array<string|int, mixed>getReactions()
public
getReactions() : array<string|int, mixed>
Return values
array<string|int, mixed>isEntryEnd()
Tells whether a line closes a KEGG record. KEGG closes on three slashes where most flat files use two.
public
static isEntryEnd(string $sLine) : bool
Parameters
- $sLine : string
-
The line to analyze
Return values
boolisEntryStart()
Tells whether a line opens a new KEGG record.
public
static isEntryStart(string $sLine) : bool
Parameters
- $sLine : string
-
The line to analyze
Return values
boolparseDataFile()
Parses a KEGG compound data file and populates this manager's fields.
public
parseDataFile(array<string|int, mixed> $aFlines) : Sequence
Parameters
- $aFlines : array<string|int, mixed>
-
The lines the script has to parse
Tags
Return values
Sequence —$oSequence
joinLines()
Joins the lines of a field into one string. A line opening with "$" continues the word the line above broke off, so it joins without a space.
protected
joinLines(array<string|int, mixed> $aLines) : string
Parameters
- $aLines : array<string|int, mixed>
Return values
stringparseDbLinks()
Reads a DBLINKS field into pairs of database name and identifier, one per line.
protected
parseDbLinks(array<string|int, mixed> $aLines) : array<string|int, mixed>
Format : DBLINKS CAS: 50-99-7
Parameters
- $aLines : array<string|int, mixed>
Return values
array<string|int, mixed>parsePathways()
Reads a PATHWAY field into pairs of map identifier and pathway name.
protected
parsePathways(array<string|int, mixed> $aLines) : array<string|int, mixed>
Format : PATHWAY PATH: map00010 Glycolysis / Gluconeogenesis
Parameters
- $aLines : array<string|int, mixed>
Return values
array<string|int, mixed>readData()
Reads the data a line carries, which starts at its thirteenth column.
protected
static readData(string $sLine) : string
Parameters
- $sLine : string
-
The line to analyze
Return values
stringreadEntryId()
Reads the identifier out of an ENTRY line. Its last word names the kind of record rather than the record itself - "C00031 Compound", "EC 2.7.1.1 Enzyme" - so the identifier is what comes before it, which for an enzyme is the two words "EC" and its number.
protected
static readEntryId(string $sData) : string
Parameters
- $sData : string
Return values
stringreadFields()
Gathers a record into its fields : one entry per label, holding the data of its own line and of every continuation line below it.
protected
readFields(array<string|int, mixed> $aFlines) : array<string|int, mixed>
Parameters
- $aFlines : array<string|int, mixed>
-
The lines the script has to parse
Return values
array<string|int, mixed>readLabel()
Reads the label a line carries in its first twelve columns.
protected
static readLabel(string $sLine) : string
Parameters
- $sLine : string
-
The line to analyze
Return values
stringsplitTokens()
Splits the lines of a field into the whitespace-separated identifiers they list.
protected
splitTokens(array<string|int, mixed> $aLines) : array<string|int, mixed>
Parameters
- $aLines : array<string|int, mixed>