SlideShare a Scribd company logo
1 of 20
Download to read offline
Building a Spanish MMTx by
using Automatic Translation and
Biomedical Ontologies
Francisco Carrero 1,2 ; José Carlos Cortizo 1,2 ; José Mª Gómez 3
              1    Wipley, Social Gaming Platform
                   http://www.wipley.com
               2   Universidad Europea de Madrid
                   http://www.esp.uem.es/gsi
              3    Optenet
                   http://www.esp.uem.es/gsi
Outline

   The MIRCAT project
   The challenge
   English MetaMap, a big effort
   Approaching a Spanish MetaMap
   Experiments
   Discussion of the Results and Future Work
                                               Francisco Carrero Garcia
The MIRCAT Project
The Interface




                     Francisco Carrero Garcia
The MIRCAT Project
System’s Architecture




                        Francisco Carrero Garcia
The Challenge
Our Goal




                            English docs




           Medical record


                            Spanish docs

                                           Francisco Carrero Garcia
The Challenge
The problem




     We can extract UMLS concepts from English texts using
     MetaMap...
     ...but there is no Spanish version of MetaMap
     Is it difficult to construct a tool like MetaMap?


                                                        Francisco Carrero Garcia
English MetaMap
A big Effort




                  ∼3 years!!

                        Francisco Carrero Garcia
Approaching Spanish MetaMap
Two Main Approaches Considered




                                 Francisco Carrero Garcia
Approaching Spanish MetaMap
Our Approach: Translation and Reuse




                    Optional



                                      Francisco Carrero Garcia
Experimental Design
Text Collections


      MedLine Plus medical News
          http://www.nlm.nih.gov/medlineplus/newsbydate.html
          Excellent online resource
          2000 news, some in English, some in Spanish
          600 available in both languages

                                                        Francisco Carrero Garcia
Experiments
Experimental Design

     MetaMap extracts concepts, allowing multiple representations
         A => Using compound concepts
         B => simple concepts
         1 => resolves ambiguity by adding all the concepts
         2 => ignores ambiguities by choosing the first possibility
         4 representations: A1, A2, B1, B2
                                                       Francisco Carrero Garcia
Experiments
Filtering


      Data representations containing a lot of features do not usually
      perform very well in text tasks
      Many classifiers degrade in prediction accuracy when faced with
      many irrelevant features or redundant/correlated ones (“curse
      of dimensionality”)
      We apply Zipf’s Law to filter the attributes

                                                        Francisco Carrero Garcia
Experiments Results
Number of concepts for each representation




                                             Francisco Carrero Garcia
Experiments Results
Average Similarities




                       Francisco Carrero Garcia
Experiments Results
Last Experiments (not in IDEAL paper)




                                        Francisco Carrero Garcia
Discussion of the Results
Translation

      The worst results (similarity) are achieved with the most
      complex (near to humans) representation: A1
      B1 is less complex and produces the best results
      => Our model seems to be more suitable as a plain bag-of-
      concepts representation
         Similar to bag-of-words representation, widely used in text
         processing tasks
                                                         Francisco Carrero Garcia
Discussion of the Results
Classification


      All results are comparable to classification on original English
      texts
      In some cases, are even better
      Best results using A2+Zipf, +7.8% in AUC
      UNMKD representations never achieves worse classifications than
      English

                                                         Francisco Carrero Garcia
Conclussions and Future Work

   The “easy way” to construct a Spanish MetaMap is promising
   Google Translation seems a good tool to adapt English resources
   to any other languages (like Spanish)
   We should try other translation tools
   We are working on applying this approach to other text tasks
   (like Information Retrieval and Filtering)

                                                    Francisco Carrero Garcia
Ending...




   Thank you very much for your attention




                                            Francisco Carrero Garcia
Any Question?




                Francisco Carrero Garcia

More Related Content

More from Jose Carlos Cortizo Perez

13+2 Herramientas eCommerce españolas para Vender Más
13+2 Herramientas eCommerce españolas para Vender Más13+2 Herramientas eCommerce españolas para Vender Más
13+2 Herramientas eCommerce españolas para Vender MásJose Carlos Cortizo Perez
 
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongress
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongressEcommerce B2B - Una Nueva Esperanza #B2BSalesCongress
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongressJose Carlos Cortizo Perez
 
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerce
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerceAdobe compra Magento: El sentimiento de la Comunidad Magento y eCommerce
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerceJose Carlos Cortizo Perez
 
Introducción del Visual Commerce Day #VCD18
Introducción del Visual Commerce Day #VCD18Introducción del Visual Commerce Day #VCD18
Introducción del Visual Commerce Day #VCD18Jose Carlos Cortizo Perez
 
La psicología de la Compra - Sobre Neandertales perdidos en Internet
La psicología de la Compra - Sobre Neandertales perdidos en InternetLa psicología de la Compra - Sobre Neandertales perdidos en Internet
La psicología de la Compra - Sobre Neandertales perdidos en InternetJose Carlos Cortizo Perez
 
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplify
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplifyBye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplify
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplifyJose Carlos Cortizo Perez
 
Los retos a nivel de negocio del eCommerce B2B
Los retos a nivel de negocio del eCommerce B2BLos retos a nivel de negocio del eCommerce B2B
Los retos a nivel de negocio del eCommerce B2BJose Carlos Cortizo Perez
 
The Reality of Gamified Loyalty in eCommerce - GWC2014
The Reality of Gamified Loyalty in eCommerce - GWC2014The Reality of Gamified Loyalty in eCommerce - GWC2014
The Reality of Gamified Loyalty in eCommerce - GWC2014Jose Carlos Cortizo Perez
 
Hablando de Gamificación en Botanic Fridays
Hablando de Gamificación en Botanic FridaysHablando de Gamificación en Botanic Fridays
Hablando de Gamificación en Botanic FridaysJose Carlos Cortizo Perez
 
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerce
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerceCómo la Gamificación ayuda al Funnel de Venta en #eCommerce
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerceJose Carlos Cortizo Perez
 
Convierte a tus usuarios en clientes - MesComercio 2012
Convierte a tus usuarios en clientes - MesComercio 2012Convierte a tus usuarios en clientes - MesComercio 2012
Convierte a tus usuarios en clientes - MesComercio 2012Jose Carlos Cortizo Perez
 
Redes Sociales y Videojuegos: una unión perfecta
Redes Sociales y Videojuegos: una unión perfectaRedes Sociales y Videojuegos: una unión perfecta
Redes Sociales y Videojuegos: una unión perfectaJose Carlos Cortizo Perez
 
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...Jose Carlos Cortizo Perez
 

More from Jose Carlos Cortizo Perez (20)

13+2 Herramientas eCommerce españolas para Vender Más
13+2 Herramientas eCommerce españolas para Vender Más13+2 Herramientas eCommerce españolas para Vender Más
13+2 Herramientas eCommerce españolas para Vender Más
 
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongress
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongressEcommerce B2B - Una Nueva Esperanza #B2BSalesCongress
Ecommerce B2B - Una Nueva Esperanza #B2BSalesCongress
 
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerce
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerceAdobe compra Magento: El sentimiento de la Comunidad Magento y eCommerce
Adobe compra Magento: El sentimiento de la Comunidad Magento y eCommerce
 
Introducción del Visual Commerce Day #VCD18
Introducción del Visual Commerce Day #VCD18Introducción del Visual Commerce Day #VCD18
Introducción del Visual Commerce Day #VCD18
 
La psicología de la Compra - Sobre Neandertales perdidos en Internet
La psicología de la Compra - Sobre Neandertales perdidos en InternetLa psicología de la Compra - Sobre Neandertales perdidos en Internet
La psicología de la Compra - Sobre Neandertales perdidos en Internet
 
Fidelizacion Ecommerce: La Última Frontera
Fidelizacion Ecommerce: La Última FronteraFidelizacion Ecommerce: La Última Frontera
Fidelizacion Ecommerce: La Última Frontera
 
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplify
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplifyBye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplify
Bye Bye Personalización: La Era de las Experiencias Personales #MagnoliaAmplify
 
Black Friday 2016: ¿Qué podemos esperar?
Black Friday 2016: ¿Qué podemos esperar?Black Friday 2016: ¿Qué podemos esperar?
Black Friday 2016: ¿Qué podemos esperar?
 
Los retos a nivel de negocio del eCommerce B2B
Los retos a nivel de negocio del eCommerce B2BLos retos a nivel de negocio del eCommerce B2B
Los retos a nivel de negocio del eCommerce B2B
 
Growth Hackeando tu eCommerce
Growth Hackeando tu eCommerceGrowth Hackeando tu eCommerce
Growth Hackeando tu eCommerce
 
Gamification workshop at the QSP Summit
Gamification workshop at the QSP SummitGamification workshop at the QSP Summit
Gamification workshop at the QSP Summit
 
The Reality of Gamified Loyalty in eCommerce - GWC2014
The Reality of Gamified Loyalty in eCommerce - GWC2014The Reality of Gamified Loyalty in eCommerce - GWC2014
The Reality of Gamified Loyalty in eCommerce - GWC2014
 
Hablando de Gamificación en Botanic Fridays
Hablando de Gamificación en Botanic FridaysHablando de Gamificación en Botanic Fridays
Hablando de Gamificación en Botanic Fridays
 
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerce
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerceCómo la Gamificación ayuda al Funnel de Venta en #eCommerce
Cómo la Gamificación ayuda al Funnel de Venta en #eCommerce
 
Introducción a la Gamificación
Introducción a la GamificaciónIntroducción a la Gamificación
Introducción a la Gamificación
 
Convierte a tus usuarios en clientes - MesComercio 2012
Convierte a tus usuarios en clientes - MesComercio 2012Convierte a tus usuarios en clientes - MesComercio 2012
Convierte a tus usuarios en clientes - MesComercio 2012
 
Open Source en Educación
Open Source en EducaciónOpen Source en Educación
Open Source en Educación
 
Redes Sociales y Videojuegos: una unión perfecta
Redes Sociales y Videojuegos: una unión perfectaRedes Sociales y Videojuegos: una unión perfecta
Redes Sociales y Videojuegos: una unión perfecta
 
Emprendiendo desde la Universidad
Emprendiendo desde la UniversidadEmprendiendo desde la Universidad
Emprendiendo desde la Universidad
 
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...
Gamificacion y Docencia: o que la Universidad tiene que aprender de los Video...
 

Recently uploaded

Histor y of HAM Radio presentation slide
Histor y of HAM Radio presentation slideHistor y of HAM Radio presentation slide
Histor y of HAM Radio presentation slidevu2urc
 
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...HostedbyConfluent
 
Data Cloud, More than a CDP by Matt Robison
Data Cloud, More than a CDP by Matt RobisonData Cloud, More than a CDP by Matt Robison
Data Cloud, More than a CDP by Matt RobisonAnna Loughnan Colquhoun
 
Google AI Hackathon: LLM based Evaluator for RAG
Google AI Hackathon: LLM based Evaluator for RAGGoogle AI Hackathon: LLM based Evaluator for RAG
Google AI Hackathon: LLM based Evaluator for RAGSujit Pal
 
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure service
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure serviceWhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure service
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure servicePooja Nehwal
 
🐬 The future of MySQL is Postgres 🐘
🐬  The future of MySQL is Postgres   🐘🐬  The future of MySQL is Postgres   🐘
🐬 The future of MySQL is Postgres 🐘RTylerCroy
 
08448380779 Call Girls In Civil Lines Women Seeking Men
08448380779 Call Girls In Civil Lines Women Seeking Men08448380779 Call Girls In Civil Lines Women Seeking Men
08448380779 Call Girls In Civil Lines Women Seeking MenDelhi Call girls
 
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...Drew Madelung
 
08448380779 Call Girls In Greater Kailash - I Women Seeking Men
08448380779 Call Girls In Greater Kailash - I Women Seeking Men08448380779 Call Girls In Greater Kailash - I Women Seeking Men
08448380779 Call Girls In Greater Kailash - I Women Seeking MenDelhi Call girls
 
How to Troubleshoot Apps for the Modern Connected Worker
How to Troubleshoot Apps for the Modern Connected WorkerHow to Troubleshoot Apps for the Modern Connected Worker
How to Troubleshoot Apps for the Modern Connected WorkerThousandEyes
 
The 7 Things I Know About Cyber Security After 25 Years | April 2024
The 7 Things I Know About Cyber Security After 25 Years | April 2024The 7 Things I Know About Cyber Security After 25 Years | April 2024
The 7 Things I Know About Cyber Security After 25 Years | April 2024Rafal Los
 
Injustice - Developers Among Us (SciFiDevCon 2024)
Injustice - Developers Among Us (SciFiDevCon 2024)Injustice - Developers Among Us (SciFiDevCon 2024)
Injustice - Developers Among Us (SciFiDevCon 2024)Allon Mureinik
 
Slack Application Development 101 Slides
Slack Application Development 101 SlidesSlack Application Development 101 Slides
Slack Application Development 101 Slidespraypatel2
 
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...shyamraj55
 
Enhancing Worker Digital Experience: A Hands-on Workshop for Partners
Enhancing Worker Digital Experience: A Hands-on Workshop for PartnersEnhancing Worker Digital Experience: A Hands-on Workshop for Partners
Enhancing Worker Digital Experience: A Hands-on Workshop for PartnersThousandEyes
 
Handwritten Text Recognition for manuscripts and early printed texts
Handwritten Text Recognition for manuscripts and early printed textsHandwritten Text Recognition for manuscripts and early printed texts
Handwritten Text Recognition for manuscripts and early printed textsMaria Levchenko
 
SQL Database Design For Developers at php[tek] 2024
SQL Database Design For Developers at php[tek] 2024SQL Database Design For Developers at php[tek] 2024
SQL Database Design For Developers at php[tek] 2024Scott Keck-Warren
 
Boost PC performance: How more available memory can improve productivity
Boost PC performance: How more available memory can improve productivityBoost PC performance: How more available memory can improve productivity
Boost PC performance: How more available memory can improve productivityPrincipled Technologies
 
Salesforce Community Group Quito, Salesforce 101
Salesforce Community Group Quito, Salesforce 101Salesforce Community Group Quito, Salesforce 101
Salesforce Community Group Quito, Salesforce 101Paola De la Torre
 
GenCyber Cyber Security Day Presentation
GenCyber Cyber Security Day PresentationGenCyber Cyber Security Day Presentation
GenCyber Cyber Security Day PresentationMichael W. Hawkins
 

Recently uploaded (20)

Histor y of HAM Radio presentation slide
Histor y of HAM Radio presentation slideHistor y of HAM Radio presentation slide
Histor y of HAM Radio presentation slide
 
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...
Transforming Data Streams with Kafka Connect: An Introduction to Single Messa...
 
Data Cloud, More than a CDP by Matt Robison
Data Cloud, More than a CDP by Matt RobisonData Cloud, More than a CDP by Matt Robison
Data Cloud, More than a CDP by Matt Robison
 
Google AI Hackathon: LLM based Evaluator for RAG
Google AI Hackathon: LLM based Evaluator for RAGGoogle AI Hackathon: LLM based Evaluator for RAG
Google AI Hackathon: LLM based Evaluator for RAG
 
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure service
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure serviceWhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure service
WhatsApp 9892124323 ✓Call Girls In Kalyan ( Mumbai ) secure service
 
🐬 The future of MySQL is Postgres 🐘
🐬  The future of MySQL is Postgres   🐘🐬  The future of MySQL is Postgres   🐘
🐬 The future of MySQL is Postgres 🐘
 
08448380779 Call Girls In Civil Lines Women Seeking Men
08448380779 Call Girls In Civil Lines Women Seeking Men08448380779 Call Girls In Civil Lines Women Seeking Men
08448380779 Call Girls In Civil Lines Women Seeking Men
 
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...
Strategies for Unlocking Knowledge Management in Microsoft 365 in the Copilot...
 
08448380779 Call Girls In Greater Kailash - I Women Seeking Men
08448380779 Call Girls In Greater Kailash - I Women Seeking Men08448380779 Call Girls In Greater Kailash - I Women Seeking Men
08448380779 Call Girls In Greater Kailash - I Women Seeking Men
 
How to Troubleshoot Apps for the Modern Connected Worker
How to Troubleshoot Apps for the Modern Connected WorkerHow to Troubleshoot Apps for the Modern Connected Worker
How to Troubleshoot Apps for the Modern Connected Worker
 
The 7 Things I Know About Cyber Security After 25 Years | April 2024
The 7 Things I Know About Cyber Security After 25 Years | April 2024The 7 Things I Know About Cyber Security After 25 Years | April 2024
The 7 Things I Know About Cyber Security After 25 Years | April 2024
 
Injustice - Developers Among Us (SciFiDevCon 2024)
Injustice - Developers Among Us (SciFiDevCon 2024)Injustice - Developers Among Us (SciFiDevCon 2024)
Injustice - Developers Among Us (SciFiDevCon 2024)
 
Slack Application Development 101 Slides
Slack Application Development 101 SlidesSlack Application Development 101 Slides
Slack Application Development 101 Slides
 
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...
Automating Business Process via MuleSoft Composer | Bangalore MuleSoft Meetup...
 
Enhancing Worker Digital Experience: A Hands-on Workshop for Partners
Enhancing Worker Digital Experience: A Hands-on Workshop for PartnersEnhancing Worker Digital Experience: A Hands-on Workshop for Partners
Enhancing Worker Digital Experience: A Hands-on Workshop for Partners
 
Handwritten Text Recognition for manuscripts and early printed texts
Handwritten Text Recognition for manuscripts and early printed textsHandwritten Text Recognition for manuscripts and early printed texts
Handwritten Text Recognition for manuscripts and early printed texts
 
SQL Database Design For Developers at php[tek] 2024
SQL Database Design For Developers at php[tek] 2024SQL Database Design For Developers at php[tek] 2024
SQL Database Design For Developers at php[tek] 2024
 
Boost PC performance: How more available memory can improve productivity
Boost PC performance: How more available memory can improve productivityBoost PC performance: How more available memory can improve productivity
Boost PC performance: How more available memory can improve productivity
 
Salesforce Community Group Quito, Salesforce 101
Salesforce Community Group Quito, Salesforce 101Salesforce Community Group Quito, Salesforce 101
Salesforce Community Group Quito, Salesforce 101
 
GenCyber Cyber Security Day Presentation
GenCyber Cyber Security Day PresentationGenCyber Cyber Security Day Presentation
GenCyber Cyber Security Day Presentation
 

Presentación en IDEAL 2008

  • 1. Building a Spanish MMTx by using Automatic Translation and Biomedical Ontologies Francisco Carrero 1,2 ; José Carlos Cortizo 1,2 ; José Mª Gómez 3 1 Wipley, Social Gaming Platform http://www.wipley.com 2 Universidad Europea de Madrid http://www.esp.uem.es/gsi 3 Optenet http://www.esp.uem.es/gsi
  • 2. Outline The MIRCAT project The challenge English MetaMap, a big effort Approaching a Spanish MetaMap Experiments Discussion of the Results and Future Work Francisco Carrero Garcia
  • 3. The MIRCAT Project The Interface Francisco Carrero Garcia
  • 4. The MIRCAT Project System’s Architecture Francisco Carrero Garcia
  • 5. The Challenge Our Goal English docs Medical record Spanish docs Francisco Carrero Garcia
  • 6. The Challenge The problem We can extract UMLS concepts from English texts using MetaMap... ...but there is no Spanish version of MetaMap Is it difficult to construct a tool like MetaMap? Francisco Carrero Garcia
  • 7. English MetaMap A big Effort ∼3 years!! Francisco Carrero Garcia
  • 8. Approaching Spanish MetaMap Two Main Approaches Considered Francisco Carrero Garcia
  • 9. Approaching Spanish MetaMap Our Approach: Translation and Reuse Optional Francisco Carrero Garcia
  • 10. Experimental Design Text Collections MedLine Plus medical News http://www.nlm.nih.gov/medlineplus/newsbydate.html Excellent online resource 2000 news, some in English, some in Spanish 600 available in both languages Francisco Carrero Garcia
  • 11. Experiments Experimental Design MetaMap extracts concepts, allowing multiple representations A => Using compound concepts B => simple concepts 1 => resolves ambiguity by adding all the concepts 2 => ignores ambiguities by choosing the first possibility 4 representations: A1, A2, B1, B2 Francisco Carrero Garcia
  • 12. Experiments Filtering Data representations containing a lot of features do not usually perform very well in text tasks Many classifiers degrade in prediction accuracy when faced with many irrelevant features or redundant/correlated ones (“curse of dimensionality”) We apply Zipf’s Law to filter the attributes Francisco Carrero Garcia
  • 13. Experiments Results Number of concepts for each representation Francisco Carrero Garcia
  • 14. Experiments Results Average Similarities Francisco Carrero Garcia
  • 15. Experiments Results Last Experiments (not in IDEAL paper) Francisco Carrero Garcia
  • 16. Discussion of the Results Translation The worst results (similarity) are achieved with the most complex (near to humans) representation: A1 B1 is less complex and produces the best results => Our model seems to be more suitable as a plain bag-of- concepts representation Similar to bag-of-words representation, widely used in text processing tasks Francisco Carrero Garcia
  • 17. Discussion of the Results Classification All results are comparable to classification on original English texts In some cases, are even better Best results using A2+Zipf, +7.8% in AUC UNMKD representations never achieves worse classifications than English Francisco Carrero Garcia
  • 18. Conclussions and Future Work The “easy way” to construct a Spanish MetaMap is promising Google Translation seems a good tool to adapt English resources to any other languages (like Spanish) We should try other translation tools We are working on applying this approach to other text tasks (like Information Retrieval and Filtering) Francisco Carrero Garcia
  • 19. Ending... Thank you very much for your attention Francisco Carrero Garcia
  • 20. Any Question? Francisco Carrero Garcia