IQ tests are not for machines, yet
[EN] Complex, but specific, tasks¿such as chess or Jeopardy!¿are popularly seen as milestones for artificial intelligence (AI). However, they are not appropriate for evaluating the intelligence of machines or measuring the progress in AI. Aware of this delusion, Detterman has recently raised a chall...
| Autores: | , |
|---|---|
| Tipo de recurso: | artículo |
| Fecha de publicación: | 2012 |
| País: | España |
| Institución: | Universitat Politècnica de València (UPV) |
| Repositorio: | RiuNet. Repositorio Institucional de la Universitat Politécnica de Valéncia |
| Idioma: | inglés |
| OAI Identifier: | oai:riunet.upv.es:10251/36157 |
| Acceso en línea: | https://riunet.upv.es/handle/10251/36157 |
| Access Level: | acceso abierto |
| Palabra clave: | Machine intelligence evaluation IQ tests Artificial intelligence Universal tests Psychometrics Task difficulty CAPTCHA LENGUAJES Y SISTEMAS INFORMATICOS |
| id |
ES_e265a9267dde33cae8b5b87e4ef8ef61 |
|---|---|
| oai_identifier_str |
oai:riunet.upv.es:10251/36157 |
| network_acronym_str |
ES |
| network_name_str |
España |
| repository_id_str |
|
| spelling |
IQ tests are not for machines, yetDowe, David L.Hernández-Orallo, José|||0000-0001-9746-7632Machine intelligence evaluationIQ testsArtificial intelligenceUniversal testsPsychometricsTask difficultyCAPTCHALENGUAJES Y SISTEMAS INFORMATICOS[EN] Complex, but specific, tasks¿such as chess or Jeopardy!¿are popularly seen as milestones for artificial intelligence (AI). However, they are not appropriate for evaluating the intelligence of machines or measuring the progress in AI. Aware of this delusion, Detterman has recently raised a challenge prompting AI researchers to evaluate their artefacts against IQ tests. We agree that the philosophy behind (human) IQ tests is a much better approach to machine intelligence evaluation than these specific tasks, and also more practical and informative than the Turing test. However, we have first to recall some work on machine intelligence measurement which has shown that some IQ tests can be passed by relatively simple programs. This suggests that the challenge may not be so demanding and may just work as a sophisticated CAPTCHA, since some types of tests might be easier than others for the current state of AI. Second, we show that an alternative, formal derivation of intelligence tests for machines is possible, grounded in (algorithmic) information theory. In these tests, we have a proper mathematical definition of what is being measured. Third, we re-visit some research done in the past fifteen years for effectively measuring machine intelligence¿since some assumptions about the subjects and their distribution no longer hold.This work was supported by the MEC projects EXPLORAINGENIO TIN 2009-06078-E, CONSOLIDER-INGENIO 26706 and TIN 2010-21062-C02-02, and GVA project PROMETEO/2008/051.ElsevierDepartamento de Sistemas Informáticos y ComputaciónEscuela Técnica Superior de Ingeniería InformáticaInstituto Universitario Valenciano de Investigación en Inteligencia ArtificialGeneralitat ValencianaMinisterio de Ciencia e InnovaciónRepositorio Institucional de la Universitat Politècnica de València Riunet20122012-04-01journal articlehttp://purl.org/coar/resource_type/c_6501VoRhttp://purl.org/coar/version/c_970fb48d4fbd8a85info:eu-repo/semantics/articleapplication/pdfapplication/pdfhttps://riunet.upv.es/handle/10251/36157reponame:RiuNet. Repositorio Institucional de la Universitat Politécnica de Valénciainstname:Universitat Politècnica de València (UPV)InglésengMinisterio de Ciencia e Innovación http://dx.doi.org/10.13039/501100004837 TIN2009-06078-E ANYTIME UNIVERSAL INTELLIGENCEMinisterio de Ciencia e Innovación http://dx.doi.org/10.13039/501100004837 TIN2010-21062-C02-02 SWEETLOGICS-UPVGeneralitat Valenciana https://doi.org/10.13039/501100003359 PROMETEO08%2F2008%2F051 Advances on Agreement Technologies for Computational Entities (atforce)open accesshttp://purl.org/coar/access_right/c_abf2Reserva de todos los derechoshttp://rightsstatements.org/vocab/InC/1.0/info:eu-repo/semantics/openAccessoai:riunet.upv.es:10251/361572026-06-13T07:49:27Z |
| dc.title.none.fl_str_mv |
IQ tests are not for machines, yet |
| title |
IQ tests are not for machines, yet |
| spellingShingle |
IQ tests are not for machines, yet Dowe, David L. Machine intelligence evaluation IQ tests Artificial intelligence Universal tests Psychometrics Task difficulty CAPTCHA LENGUAJES Y SISTEMAS INFORMATICOS |
| title_short |
IQ tests are not for machines, yet |
| title_full |
IQ tests are not for machines, yet |
| title_fullStr |
IQ tests are not for machines, yet |
| title_full_unstemmed |
IQ tests are not for machines, yet |
| title_sort |
IQ tests are not for machines, yet |
| dc.creator.none.fl_str_mv |
Dowe, David L. Hernández-Orallo, José|||0000-0001-9746-7632 |
| author |
Dowe, David L. |
| author_facet |
Dowe, David L. Hernández-Orallo, José|||0000-0001-9746-7632 |
| author_role |
author |
| author2 |
Hernández-Orallo, José|||0000-0001-9746-7632 |
| author2_role |
author |
| dc.contributor.none.fl_str_mv |
Departamento de Sistemas Informáticos y Computación Escuela Técnica Superior de Ingeniería Informática Instituto Universitario Valenciano de Investigación en Inteligencia Artificial Generalitat Valenciana Ministerio de Ciencia e Innovación Repositorio Institucional de la Universitat Politècnica de València Riunet |
| dc.subject.none.fl_str_mv |
Machine intelligence evaluation IQ tests Artificial intelligence Universal tests Psychometrics Task difficulty CAPTCHA LENGUAJES Y SISTEMAS INFORMATICOS |
| topic |
Machine intelligence evaluation IQ tests Artificial intelligence Universal tests Psychometrics Task difficulty CAPTCHA LENGUAJES Y SISTEMAS INFORMATICOS |
| description |
[EN] Complex, but specific, tasks¿such as chess or Jeopardy!¿are popularly seen as milestones for artificial intelligence (AI). However, they are not appropriate for evaluating the intelligence of machines or measuring the progress in AI. Aware of this delusion, Detterman has recently raised a challenge prompting AI researchers to evaluate their artefacts against IQ tests. We agree that the philosophy behind (human) IQ tests is a much better approach to machine intelligence evaluation than these specific tasks, and also more practical and informative than the Turing test. However, we have first to recall some work on machine intelligence measurement which has shown that some IQ tests can be passed by relatively simple programs. This suggests that the challenge may not be so demanding and may just work as a sophisticated CAPTCHA, since some types of tests might be easier than others for the current state of AI. Second, we show that an alternative, formal derivation of intelligence tests for machines is possible, grounded in (algorithmic) information theory. In these tests, we have a proper mathematical definition of what is being measured. Third, we re-visit some research done in the past fifteen years for effectively measuring machine intelligence¿since some assumptions about the subjects and their distribution no longer hold. |
| publishDate |
2012 |
| dc.date.none.fl_str_mv |
2012 2012-04-01 |
| dc.type.none.fl_str_mv |
journal article http://purl.org/coar/resource_type/c_6501 VoR http://purl.org/coar/version/c_970fb48d4fbd8a85 |
| dc.type.openaire.fl_str_mv |
info:eu-repo/semantics/article |
| format |
article |
| dc.identifier.none.fl_str_mv |
https://riunet.upv.es/handle/10251/36157 |
| url |
https://riunet.upv.es/handle/10251/36157 |
| dc.language.none.fl_str_mv |
Inglés eng |
| language_invalid_str_mv |
Inglés |
| language |
eng |
| dc.relation.none.fl_str_mv |
Ministerio de Ciencia e Innovación http://dx.doi.org/10.13039/501100004837 TIN2009-06078-E ANYTIME UNIVERSAL INTELLIGENCE Ministerio de Ciencia e Innovación http://dx.doi.org/10.13039/501100004837 TIN2010-21062-C02-02 SWEETLOGICS-UPV Generalitat Valenciana https://doi.org/10.13039/501100003359 PROMETEO08%2F2008%2F051 Advances on Agreement Technologies for Computational Entities (atforce) |
| dc.rights.none.fl_str_mv |
open access http://purl.org/coar/access_right/c_abf2 Reserva de todos los derechos http://rightsstatements.org/vocab/InC/1.0/ |
| dc.rights.openaire.fl_str_mv |
info:eu-repo/semantics/openAccess |
| rights_invalid_str_mv |
open access http://purl.org/coar/access_right/c_abf2 Reserva de todos los derechos http://rightsstatements.org/vocab/InC/1.0/ |
| eu_rights_str_mv |
openAccess |
| dc.format.none.fl_str_mv |
application/pdf application/pdf |
| dc.publisher.none.fl_str_mv |
Elsevier |
| publisher.none.fl_str_mv |
Elsevier |
| dc.source.none.fl_str_mv |
reponame:RiuNet. Repositorio Institucional de la Universitat Politécnica de Valéncia instname:Universitat Politècnica de València (UPV) |
| instname_str |
Universitat Politècnica de València (UPV) |
| reponame_str |
RiuNet. Repositorio Institucional de la Universitat Politécnica de Valéncia |
| collection |
RiuNet. Repositorio Institucional de la Universitat Politécnica de Valéncia |
| repository.name.fl_str_mv |
|
| repository.mail.fl_str_mv |
|
| _version_ |
1869422381978615808 |
| score |
15.301603 |