![]()  | 
ISCApad Archive  »  2010  »  ISCApad #149  »  Resources  »  Database  »  ELRA Language Resources Catalogue Update (June 2010)  | 
ISCApad #149 | 
| Friday, November 05, 2010 by Chris Wellekens | 
ELRA is happy to announce that 2 new Speech Desktop/Microphone resources, 1 new Terminological Resource and 1 Written Corpus are now available in its catalogue: ELRA-S0305  EPAC Corpus: orthographic transcriptions 
This corpus consists  of approx. 100 hours of manual orthographic transcriptions, which were  produced from 1,677 hours of non transcribed recordings from the ESTER  Evaluation Campaign (Technolangue programme). This corpus also consists  of automatic transcriptions of the full 1,677 hours. 
For more  information, see:  http://catalog.elra.info/product_info.php?products_id=1119 
ELRA-S0307  BABEL Polish database 
The BABEL Polish Database is a speech  database that was produced by a research consortium funded by the  European Union under the COPERNICUS programme (COPERNICUS Project 1304).  It consists of the basic 'common' set which contains the Many Talker  Set (30 males, 30 females), the Few Talker Set (5 males, 5 females), the  Very Few Talker Set (1 male, 1 female). 
For more information,  see: http://catalog.elra.info/product_info.php?products_id=1120 
ELRA-T0374  Terminology database of natural sciences 
This dictionary  covers the three kingdoms: Animal, Vegetal, Mineral. It contains 50,000  species with numerous synonyms in French, English and Latin and many  breeds and varieties. Minerals are given with their chemical formula.  About 7,900 definitions in French are included. It also includes  synonyms and linguistic variants. 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=1121 
ELRA-W0053  Catalan-Spanish Parallel Corpus 
This corpus contains more  than 100 million words and it contains 10 years of bilingual articles  from “El Periódico de Catalunya”. The data are aligned at sentence level  and stored in text files, in a one sentence per line basis. The data  are provided in plain text, with no encoding whatsoever. 
For  more information, see:  http://catalog.elra.info/product_info.php?products_id=1122 
****** 
Moreover,  please note that the content of the following 3 Terminological  Resources has been updated and their prices have been revised: 
ELRA-T0102  Terminology database of expressions 
This resource comprises  over about 26,000-30,000 expressions, such as sayings, proverbs, idioms,  slogans, citations, exclamations, onomatopoeias and figurative  expressions of French and English. Several grammatical topics that are  included in some sentences are also handled. This resource contains  synonyms. The DISCIPLINE field refers to the expression category:  proverbs, idioms, postposition verbs. 
For more information,  see: http://catalog.elra.info/product_info.php?products_id=114 
ELRA-T0103  Terminology database of finance  
This dictionary covers the  three kingdoms: Animal, Vegetal, Mineral. It contains 50,000 species  with numerous synonyms in French, English and Latin and many breeds and  varieties. Minerals are given with their chemical formula. About 7,900  definitions in French are included. It also includes synonyms and  linguistic variants. 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=115 
ELRA-T0367  Terminology database of telecommunication 
This resource  comprises over 89,200 entries in the field of telecommunication. It also  contains many synonyms and abbreviations in both languages, as well as  meaning, case or applications for polysemic terms. 
For more  information, see:  http://catalog.elra.info/product_info.php?products_id=659 
For  more information on the catalogue, please contact Valérie Mapelli  mailto:mapelli@elda.org 
Visit our On-line  Catalogue: http://catalog.elra.info 
Visit the Universal  Catalogue: http://universal.elra.info  
Archives of ELRA  Language Resources Catalogue Updates:  http://www.elra.info/LRs-Announcements.html    
*****************************************************************  
ELRA  - Language Resources Catalogue - Update  
*****************************************************************  
In  the framework of our ongoing campaign for updating and reducing the  prices of the language resources distributed in the ELRA catalogue, ELRA  is happy to announce that the prices for the following resources have  been substantially reduced: 
ELRA-S0074 British  English SpeechDat(II) MDB-1000 
This speech database contains  the recordings of 1,000 British speakers recorded over the British  mobile telephone network. Each speaker uttered around 40 read and  spontaneous items. 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=723 
ELRA-S0075  Welsh SpeechDat(II) FDB-2000 
This speech database contains  the recordings of 2,000 Welsh speakers recorded over the British fixed  telephone network. Each speaker uttered around 40 read and spontaneous  items. 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=557 
ELRA-S0101  Spanish SpeechDat(II) FDB-1000  
This speech database contains  the recordings of 1,000 Castillan Spanish speakers recorded over the  Spanish fixed telephone network. Each speaker uttered around 40 read and  spontaneous items.  
This database is a subset of the Spanish  SpeechDat(II) FDB-4000 (ref. ELRA-S0102). 
For more  information, see:  http://catalog.elra.info/product_info.php?products_id=726 
ELRA-S0102  Spanish SpeechDat(II) FDB-4000 
This speech database contains  the recordings of 4,000 Castillan Spanish speakers recorded over the  Spanish fixed telephone network. Each speaker uttered around 40 read and  spontaneous items. 
This database includes the Spanish  SpeechDat(II) FDB-1000 (ref. ELRA-S0101). 
For more  information, see:  http://catalog.elra.info/product_info.php?products_id=727 
ELRA-S0140  Spanish SpeechDat-Car database 
The Spanish SpeechDat-Car  database contains the recordings in a car of 306 speakers, who uttered  around 120 read and spontaneous items. Recordings have been made through  5 different channels, of which 4 were in-car microphones (1 close-talk  microphone, 3 far-talk microphones) and 1 channel over the GSM network. 
For  more information, see:  http://catalog.elra.info/product_info.php?products_id=690 
ELRA-S0141  SALA Spanish Venezuelan Database  
This speech database  contains the recordings of 1,000 Venezuelan speakers recorded over the  Venezuelan fixed telephone network. Each speaker uttered around 50 read  and spontaneous items. 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=736 
ELRA-S0297  Hungarian Speecon database  
The Hungarian Speecon database  comprises the recordings of 555 adult Hungarian speakers and 50 child  Hungarian speakers who uttered respectively over 290 items and 210 items  (read and spontaneous). 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=1094 
ELRA-S0298  Czech Speecon database 
The Czech Speecon database comprises  the recordings of 550 adult Czech speakers and 50 child Czech speakers  who uttered respectively over 290 items and 210 items (read and  spontaneous). 
For more information, see:  http://catalog.elra.info/product_info.php?products_id=1095 
For  more information on the catalogue, please contact Valérie Mapelli  mailto:mapelli@elda.org 
Visit our On-line  Catalogue: http://catalog.elra.info 
Visit the Universal  Catalogue: http://universal.elra.info  
Archives of ELRA  Language Resources Catalogue Updates:  http://www.elra.info/LRs-Announcements.html 
 | 
![]()  | Back | Top |