home.social

#alphazero — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #alphazero, aggregated by home.social.

  1. 5/
    The Future of the Infinity Machine

    DeepMind positions AI as a fundamental force of nature designed to uncover the discoverable patterns of the universe. While this "infinity machine" offers the potential to cure diseases and solve climate change, it also presents a central tension: the challenge of managing an intelligence that may eventually surpass human ability to understand or control it.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine

  2. 5/
    The Future of the Infinity Machine

    DeepMind positions AI as a fundamental force of nature designed to uncover the discoverable patterns of the universe. While this "infinity machine" offers the potential to cure diseases and solve climate change, it also presents a central tension: the challenge of managing an intelligence that may eventually surpass human ability to understand or control it.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine

  3. 4/
    #AlphaZero: A breakthrough agent that learned only from the rules of a game, eventually discovering strategies previously unknown to humans.

    AlphaFold: A massive pivot to biology that solved the 50-year-old "protein folding" challenge. AlphaFold2 achieved accuracy levels comparable to years of expensive lab work, allowing scientists to map antibiotic-resistant bugs and design disease-resistant crops in minutes.

    youtu.be/2nTuBwvp1Is

    #DeepMind #AI #AGI #AlphaFold #InfinityMachine

  4. 4/
    #AlphaZero: A breakthrough agent that learned only from the rules of a game, eventually discovering strategies previously unknown to humans.

    AlphaFold: A massive pivot to biology that solved the 50-year-old "protein folding" challenge. AlphaFold2 achieved accuracy levels comparable to years of expensive lab work, allowing scientists to map antibiotic-resistant bugs and design disease-resistant crops in minutes.

    youtu.be/2nTuBwvp1Is

    #DeepMind #AI #AGI #AlphaFold #InfinityMachine

  5. Founding Vision and Philosophy

    DeepMind's founder, Demis Hassabis, is a former chess prodigy who transitioned from mastering games to studying the rules of intelligence through neuroscience. His core philosophy is built on two primary beliefs:

    Intelligence as a Lens:
    To truly understand reality, one must first understand the intelligence through which we perceive it.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine #Gemini

  6. Founding Vision and Philosophy

    DeepMind's founder, Demis Hassabis, is a former chess prodigy who transitioned from mastering games to studying the rules of intelligence through neuroscience. His core philosophy is built on two primary beliefs:

    Intelligence as a Lens:
    To truly understand reality, one must first understand the intelligence through which we perceive it.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine #Gemini

  7. THE INFINITY MACHINE

    Introduction: The Quest for Super Intelligence

    The quest to build artificial general intelligence (AGI) is a modern pursuit with high stakes, echoing the awe and apprehension felt by early atomic scientists. At the center of this effort is DeepMind, a company founded on the grand vision of creating an "infinity machine"—a general-purpose tool for pure discovery.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine #Gemini

  8. THE INFINITY MACHINE

    Introduction: The Quest for Super Intelligence

    The quest to build artificial general intelligence (AGI) is a modern pursuit with high stakes, echoing the awe and apprehension felt by early atomic scientists. At the center of this effort is DeepMind, a company founded on the grand vision of creating an "infinity machine"—a general-purpose tool for pure discovery.

    youtu.be/2nTuBwvp1Is

    #DemisHassabis #DeepMind #AI #AGI #AlphaZero #AlphaFold #InfinityMachine #Gemini

  9. 13-Mar-2026
    #AI’s #gamePlaying still has flaws: #AlphaZero-style self-play tested on #Nim
    Despite heavy training, agents show blind spots and can miss optimal moves

    eurekalert.org/news-releases/1

    #science #technology

  10. 13-Mar-2026
    #AI’s #gamePlaying still has flaws: #AlphaZero-style self-play tested on #Nim
    Despite heavy training, agents show blind spots and can miss optimal moves

    eurekalert.org/news-releases/1

    #science #technology

  11. Абсолютный ноль: как ИИ учится без данных

    ​Absolute Zero Reasoner отличается от традиционных подходов к обучению ИИ, позволяя ИИ обучаться с нуля, без необходимости использования заранее предоставленных человеком данных.

    Absolute Zero Reasoner (AZR) представляет собой революционную концепцию в области искусственного...

    #DST #DSTGlobal #ДСТ #ДСТГлобал #Абсолютныйноль #искусственныйинтеллект #AbsoluteZero #ИИ #AZR #AlphaZero #DeepMind #парадигмы #Модель

    Источник: dstglobal.ru/club/1109-absolyu

  12. Абсолютный ноль: как ИИ учится без данных

    ​Absolute Zero Reasoner отличается от традиционных подходов к обучению ИИ, позволяя ИИ обучаться с нуля, без необходимости использования заранее предоставленных человеком данных.

    Absolute Zero Reasoner (AZR) представляет собой революционную концепцию в области искусственного...

    #DST #DSTGlobal #ДСТ #ДСТГлобал #Абсолютныйноль #искусственныйинтеллект #AbsoluteZero #ИИ #AZR #AlphaZero #DeepMind #парадигмы #Модель

    Источник: dstglobal.ru/club/1109-absolyu

  13. Абсолютный ноль: как ИИ учится без данных

    ​Absolute Zero Reasoner отличается от традиционных подходов к обучению ИИ, позволяя ИИ обучаться с нуля, без необходимости использования заранее предоставленных человеком данных.

    Absolute Zero Reasoner (AZR) представляет собой революционную концепцию в области искусственного...

    #DST #DSTGlobal #ДСТ #ДСТГлобал #Абсолютныйноль #искусственныйинтеллект #AbsoluteZero #ИИ #AZR #AlphaZero #DeepMind #парадигмы #Модель

    Источник: dstglobal.ru/club/1109-absolyu

  14. Абсолютный ноль: как ИИ учится без данных

    ​Absolute Zero Reasoner отличается от традиционных подходов к обучению ИИ, позволяя ИИ обучаться с нуля, без необходимости использования заранее предоставленных человеком данных.

    Absolute Zero Reasoner (AZR) представляет собой революционную концепцию в области искусственного...

    #DST #DSTGlobal #ДСТ #ДСТГлобал #Абсолютныйноль #искусственныйинтеллект #AbsoluteZero #ИИ #AZR #AlphaZero #DeepMind #парадигмы #Модель

    Источник: dstglobal.ru/club/1109-absolyu

  15. Абсолютный ноль: как ИИ учится без данных

    ​Absolute Zero Reasoner отличается от традиционных подходов к обучению ИИ, позволяя ИИ обучаться с нуля, без необходимости использования заранее предоставленных человеком данных.

    Absolute Zero Reasoner (AZR) представляет собой революционную концепцию в области искусственного...

    #DST #DSTGlobal #ДСТ #ДСТГлобал #Абсолютныйноль #искусственныйинтеллект #AbsoluteZero #ИИ #AZR #AlphaZero #DeepMind #парадигмы #Модель

    Источник: dstglobal.ru/club/1109-absolyu

  16. Oops, I think I've gone a bit too deep into the #AI rabbit hole today 😳 (a thread 🧵):

    Did you know why AI systems like #AlphaGo or #AlphaZero performed so well?
    It was because of their _objective function_:
    -1 for loosing, +1 for winning ¯\_(ツ)_/¯

    Why Artificial Intelligence Like AlphaZero Has Trouble With the Real World (February 2018)

    quantamagazine.org/why-artific

    Try to design an objective function for a self-driving car...

    1/3

    #ArtificialIntelligence #RabbitHole

  17. Oops, I think I've gone a bit too deep into the #AI rabbit hole today 😳 (a thread 🧵):

    Did you know why AI systems like #AlphaGo or #AlphaZero performed so well?
    It was because of their _objective function_:
    -1 for loosing, +1 for winning ¯\_(ツ)_/¯

    Why Artificial Intelligence Like AlphaZero Has Trouble With the Real World (February 2018)

    quantamagazine.org/why-artific

    Try to design an objective function for a self-driving car...

    1/3

    #ArtificialIntelligence #RabbitHole

  18. Pequeños y grandes pasos hacia el imperio de la inteligencia artificial

    Fuente: Open Tech

    Traducción de la infografía:

    • 1943 – McCullock y Pitts publican un artículo titulado Un cálculo lógico de ideas inmanentes en la actividad nerviosa, en el que proponen las bases para las redes neuronales.
    • 1950 – Turing publica Computing Machinery and Intelligence, proponiendo el Test de Turing como forma de medir la capacidad de una máquina.
    • 1951 – Marvin Minsky y Dean Edmonds construyen SNAR, la primera computadora de red neuronal.
    • 1956 – Se celebra la Conferencia de Dartmouth (organizada por McCarthy, Minsky, Rochester y Shannon), que marca el nacimiento de la IA como campo de estudio.
    • 1957 – Rosenblatt desarrolla el Perceptrón: la primera red neuronal artificial capaz de aprender.

    (!!) Test de Turing: donde un evaluador humano entabla una conversación en lenguaje natural con una máquina y un humano.

    • 1965 – Weizenbaum desarrolla ELIZA: un programa de procesamiento del lenguaje natural que simula una conversación.
    • 1967 – Newell y Simon desarrollan el Solucionador General de Problemas (GPS), uno de los primeros programas de IA que demuestra una capacidad de resolución de problemas similar a la humana.
    • 1974 – Comienza el primer invierno de la IA, marcado por una disminución de la financiación y del interés en la investigación en IA debido a expectativas poco realistas y a un progreso limitado.
    • 1980 – Los sistemas expertos ganan popularidad y las empresas los utilizan para realizar previsiones financieras y diagnósticos médicos.
    • 1986 – Hinton, Rumelhart y Williams publican Aprendizaje de representaciones mediante retropropagación de errores, que permite entrenar redes neuronales mucho más profundas.

    (!!) Redes neuronales: modelos de aprendizaje automático que imitan el cerebro y aprenden a reconocer patrones y hacer predicciones a través de conexiones neuronales artificiales.

    • 1997 – Deep Blue de IBM derrota al campeón mundial de ajedrez Kasparov, siendo la primera vez que una computadora vence a un campeón mundial en un juego complejo.
    • 2002 – iRobot presenta Roomba, el primer robot aspirador doméstico producido en serie con un sistema de navegación impulsado por IA.
    • 2011 – Watson de IBM derrota a dos ex campeones de Jeopardy!.
    • 2012 – La startup de inteligencia artificial DeepMind desarrolla una red neuronal profunda que puede reconocer gatos en vídeos de YouTube.
    • 2014 – Facebook crea DeepFace, un sistema de reconocimiento facial que puede reconocer rostros con una precisión casi humana.

    (!!) DeepMind fue adquirida por Google en 2014 por 500 millones de dólares.

    • 2015 – AlphaGo, desarrollado por DeepMind, derrota al campeón mundial Lee Sedol en el juego de Go.
    • 2017 – AlphaZero de Google derrota a los mejores motores de ajedrez y shogi del mundo en una serie de partidas.
    • 2020 – OpenAI lanza GPT-3, lo que marca un avance significativo en el procesamiento del lenguaje natural.

    (!!) Procesamiento del lenguaje natural: enseña a las computadoras a comprender y utilizar el lenguaje humano mediante técnicas como el aprendizaje automático.

    • 2021 – AlphaFold2 de DeepMind resuelve el problema del plegamiento de proteínas, allanando el camino para nuevos descubrimientos de fármacos y avances médicos.
    • 2022 – Google despide al ingeniero Blake Lemoine por sus afirmaciones de que el modelo de lenguaje para aplicaciones de diálogo (LaMDA) de Google era sensible.
    • 2023 – Artistas presentaron una demanda colectiva contra Stability AI, DeviantArt y Mid-journey por usar Stable Diffusion para remezclar las obras protegidas por derechos de autor de millones de artistas.

    Gráfico: Open Tech / Genuine Impact

    Entradas relacionadas

    #ajedrez #AlphaFold2 #AlphaGo #AlphaZero #aprendizajeAutomático #artículo #artistas #aspirador #BlakeLemoine #ConferenciaDeDartmouth #copyright #DeanEdmonds #DeepBlue #DeepFace #DeepMind #DeviantArt #ELIZA #Facebook #gatos #GenuineImpact #Go #Google #GPS #GPT3 #gráfico #Hinton #IA #IBM #infografía #inteligenciaArtificial #iRobot #Jeopardy_ #Kasparov #LaMDA #LeeSedol #MarvinMinsky #McCarthy #McCullock #MidJourney #modelos #Newell #OpenTech #OpenAI #patrones #Perceptron #Pitts #plegamientoDeProteínas #predicciones #procesamientoDelLenguajeNatural #reconocimientoFacial #redesNeuronales #remezclar #robot #Rochester #Roomba #Rosenblatt #Rumelhart #Shannon #shogi #Simon #sistemaDeNavegación #SNAR #StabilityAI #StableDiffusion #testDeTuring #Turing #vídeos #Watson #Weizenbaum #Williams #YouTube

  19. Pequeños y grandes pasos hacia el imperio de la inteligencia artificial

    Fuente: Open Tech

    Traducción de la infografía:

    • 1943 – McCullock y Pitts publican un artículo titulado Un cálculo lógico de ideas inmanentes en la actividad nerviosa, en el que proponen las bases para las redes neuronales.
    • 1950 – Turing publica Computing Machinery and Intelligence, proponiendo el Test de Turing como forma de medir la capacidad de una máquina.
    • 1951 – Marvin Minsky y Dean Edmonds construyen SNAR, la primera computadora de red neuronal.
    • 1956 – Se celebra la Conferencia de Dartmouth (organizada por McCarthy, Minsky, Rochester y Shannon), que marca el nacimiento de la IA como campo de estudio.
    • 1957 – Rosenblatt desarrolla el Perceptrón: la primera red neuronal artificial capaz de aprender.

    (!!) Test de Turing: donde un evaluador humano entabla una conversación en lenguaje natural con una máquina y un humano.

    • 1965 – Weizenbaum desarrolla ELIZA: un programa de procesamiento del lenguaje natural que simula una conversación.
    • 1967 – Newell y Simon desarrollan el Solucionador General de Problemas (GPS), uno de los primeros programas de IA que demuestra una capacidad de resolución de problemas similar a la humana.
    • 1974 – Comienza el primer invierno de la IA, marcado por una disminución de la financiación y del interés en la investigación en IA debido a expectativas poco realistas y a un progreso limitado.
    • 1980 – Los sistemas expertos ganan popularidad y las empresas los utilizan para realizar previsiones financieras y diagnósticos médicos.
    • 1986 – Hinton, Rumelhart y Williams publican Aprendizaje de representaciones mediante retropropagación de errores, que permite entrenar redes neuronales mucho más profundas.

    (!!) Redes neuronales: modelos de aprendizaje automático que imitan el cerebro y aprenden a reconocer patrones y hacer predicciones a través de conexiones neuronales artificiales.

    • 1997 – Deep Blue de IBM derrota al campeón mundial de ajedrez Kasparov, siendo la primera vez que una computadora vence a un campeón mundial en un juego complejo.
    • 2002 – iRobot presenta Roomba, el primer robot aspirador doméstico producido en serie con un sistema de navegación impulsado por IA.
    • 2011 – Watson de IBM derrota a dos ex campeones de Jeopardy!.
    • 2012 – La startup de inteligencia artificial DeepMind desarrolla una red neuronal profunda que puede reconocer gatos en vídeos de YouTube.
    • 2014 – Facebook crea DeepFace, un sistema de reconocimiento facial que puede reconocer rostros con una precisión casi humana.

    (!!) DeepMind fue adquirida por Google en 2014 por 500 millones de dólares.

    • 2015 – AlphaGo, desarrollado por DeepMind, derrota al campeón mundial Lee Sedol en el juego de Go.
    • 2017 – AlphaZero de Google derrota a los mejores motores de ajedrez y shogi del mundo en una serie de partidas.
    • 2020 – OpenAI lanza GPT-3, lo que marca un avance significativo en el procesamiento del lenguaje natural.

    (!!) Procesamiento del lenguaje natural: enseña a las computadoras a comprender y utilizar el lenguaje humano mediante técnicas como el aprendizaje automático.

    • 2021 – AlphaFold2 de DeepMind resuelve el problema del plegamiento de proteínas, allanando el camino para nuevos descubrimientos de fármacos y avances médicos.
    • 2022 – Google despide al ingeniero Blake Lemoine por sus afirmaciones de que el modelo de lenguaje para aplicaciones de diálogo (LaMDA) de Google era sensible.
    • 2023 – Artistas presentaron una demanda colectiva contra Stability AI, DeviantArt y Mid-journey por usar Stable Diffusion para remezclar las obras protegidas por derechos de autor de millones de artistas.

    Gráfico: Open Tech / Genuine Impact

    Entradas relacionadas

    #ajedrez #AlphaFold2 #AlphaGo #AlphaZero #aprendizajeAutomático #artículo #artistas #aspirador #BlakeLemoine #ConferenciaDeDartmouth #copyright #DeanEdmonds #DeepBlue #DeepFace #DeepMind #DeviantArt #ELIZA #Facebook #gatos #GenuineImpact #Go #Google #GPS #GPT3 #gráfico #Hinton #IA #IBM #infografía #inteligenciaArtificial #iRobot #Jeopardy_ #Kasparov #LaMDA #LeeSedol #MarvinMinsky #McCarthy #McCullock #MidJourney #modelos #Newell #OpenTech #OpenAI #patrones #Perceptron #Pitts #plegamientoDeProteínas #predicciones #procesamientoDelLenguajeNatural #reconocimientoFacial #redesNeuronales #remezclar #robot #Rochester #Roomba #Rosenblatt #Rumelhart #Shannon #shogi #Simon #sistemaDeNavegación #SNAR #StabilityAI #StableDiffusion #testDeTuring #Turing #vídeos #Watson #Weizenbaum #Williams #YouTube

  20. Pequeños y grandes pasos hacia el imperio de la inteligencia artificial

    Fuente: Open Tech

    Traducción de la infografía:

    • 1943 – McCullock y Pitts publican un artículo titulado Un cálculo lógico de ideas inmanentes en la actividad nerviosa, en el que proponen las bases para las redes neuronales.
    • 1950 – Turing publica Computing Machinery and Intelligence, proponiendo el Test de Turing como forma de medir la capacidad de una máquina.
    • 1951 – Marvin Minsky y Dean Edmonds construyen SNAR, la primera computadora de red neuronal.
    • 1956 – Se celebra la Conferencia de Dartmouth (organizada por McCarthy, Minsky, Rochester y Shannon), que marca el nacimiento de la IA como campo de estudio.
    • 1957 – Rosenblatt desarrolla el Perceptrón: la primera red neuronal artificial capaz de aprender.

    (!!) Test de Turing: donde un evaluador humano entabla una conversación en lenguaje natural con una máquina y un humano.

    • 1965 – Weizenbaum desarrolla ELIZA: un programa de procesamiento del lenguaje natural que simula una conversación.
    • 1967 – Newell y Simon desarrollan el Solucionador General de Problemas (GPS), uno de los primeros programas de IA que demuestra una capacidad de resolución de problemas similar a la humana.
    • 1974 – Comienza el primer invierno de la IA, marcado por una disminución de la financiación y del interés en la investigación en IA debido a expectativas poco realistas y a un progreso limitado.
    • 1980 – Los sistemas expertos ganan popularidad y las empresas los utilizan para realizar previsiones financieras y diagnósticos médicos.
    • 1986 – Hinton, Rumelhart y Williams publican Aprendizaje de representaciones mediante retropropagación de errores, que permite entrenar redes neuronales mucho más profundas.

    (!!) Redes neuronales: modelos de aprendizaje automático que imitan el cerebro y aprenden a reconocer patrones y hacer predicciones a través de conexiones neuronales artificiales.

    • 1997 – Deep Blue de IBM derrota al campeón mundial de ajedrez Kasparov, siendo la primera vez que una computadora vence a un campeón mundial en un juego complejo.
    • 2002 – iRobot presenta Roomba, el primer robot aspirador doméstico producido en serie con un sistema de navegación impulsado por IA.
    • 2011 – Watson de IBM derrota a dos ex campeones de Jeopardy!.
    • 2012 – La startup de inteligencia artificial DeepMind desarrolla una red neuronal profunda que puede reconocer gatos en vídeos de YouTube.
    • 2014 – Facebook crea DeepFace, un sistema de reconocimiento facial que puede reconocer rostros con una precisión casi humana.

    (!!) DeepMind fue adquirida por Google en 2014 por 500 millones de dólares.

    • 2015 – AlphaGo, desarrollado por DeepMind, derrota al campeón mundial Lee Sedol en el juego de Go.
    • 2017 – AlphaZero de Google derrota a los mejores motores de ajedrez y shogi del mundo en una serie de partidas.
    • 2020 – OpenAI lanza GPT-3, lo que marca un avance significativo en el procesamiento del lenguaje natural.

    (!!) Procesamiento del lenguaje natural: enseña a las computadoras a comprender y utilizar el lenguaje humano mediante técnicas como el aprendizaje automático.

    • 2021 – AlphaFold2 de DeepMind resuelve el problema del plegamiento de proteínas, allanando el camino para nuevos descubrimientos de fármacos y avances médicos.
    • 2022 – Google despide al ingeniero Blake Lemoine por sus afirmaciones de que el modelo de lenguaje para aplicaciones de diálogo (LaMDA) de Google era sensible.
    • 2023 – Artistas presentaron una demanda colectiva contra Stability AI, DeviantArt y Mid-journey por usar Stable Diffusion para remezclar las obras protegidas por derechos de autor de millones de artistas.

    Gráfico: Open Tech / Genuine Impact

    Entradas relacionadas

    #ajedrez #AlphaFold2 #AlphaGo #AlphaZero #aprendizajeAutomático #artículo #artistas #aspirador #BlakeLemoine #ConferenciaDeDartmouth #copyright #DeanEdmonds #DeepBlue #DeepFace #DeepMind #DeviantArt #ELIZA #Facebook #gatos #GenuineImpact #Go #Google #GPS #GPT3 #gráfico #Hinton #IA #IBM #infografía #inteligenciaArtificial #iRobot #Jeopardy_ #Kasparov #LaMDA #LeeSedol #MarvinMinsky #McCarthy #McCullock #MidJourney #modelos #Newell #OpenTech #OpenAI #patrones #Perceptron #Pitts #plegamientoDeProteínas #predicciones #procesamientoDelLenguajeNatural #reconocimientoFacial #redesNeuronales #remezclar #robot #Rochester #Roomba #Rosenblatt #Rumelhart #Shannon #shogi #Simon #sistemaDeNavegación #SNAR #StabilityAI #StableDiffusion #testDeTuring #Turing #vídeos #Watson #Weizenbaum #Williams #YouTube

  21. Pequeños y grandes pasos hacia el imperio de la inteligencia artificial

    Fuente: Open Tech

    Traducción de la infografía:

    • 1943 – McCullock y Pitts publican un artículo titulado Un cálculo lógico de ideas inmanentes en la actividad nerviosa, en el que proponen las bases para las redes neuronales.
    • 1950 – Turing publica Computing Machinery and Intelligence, proponiendo el Test de Turing como forma de medir la capacidad de una máquina.
    • 1951 – Marvin Minsky y Dean Edmonds construyen SNAR, la primera computadora de red neuronal.
    • 1956 – Se celebra la Conferencia de Dartmouth (organizada por McCarthy, Minsky, Rochester y Shannon), que marca el nacimiento de la IA como campo de estudio.
    • 1957 – Rosenblatt desarrolla el Perceptrón: la primera red neuronal artificial capaz de aprender.

    (!!) Test de Turing: donde un evaluador humano entabla una conversación en lenguaje natural con una máquina y un humano.

    • 1965 – Weizenbaum desarrolla ELIZA: un programa de procesamiento del lenguaje natural que simula una conversación.
    • 1967 – Newell y Simon desarrollan el Solucionador General de Problemas (GPS), uno de los primeros programas de IA que demuestra una capacidad de resolución de problemas similar a la humana.
    • 1974 – Comienza el primer invierno de la IA, marcado por una disminución de la financiación y del interés en la investigación en IA debido a expectativas poco realistas y a un progreso limitado.
    • 1980 – Los sistemas expertos ganan popularidad y las empresas los utilizan para realizar previsiones financieras y diagnósticos médicos.
    • 1986 – Hinton, Rumelhart y Williams publican Aprendizaje de representaciones mediante retropropagación de errores, que permite entrenar redes neuronales mucho más profundas.

    (!!) Redes neuronales: modelos de aprendizaje automático que imitan el cerebro y aprenden a reconocer patrones y hacer predicciones a través de conexiones neuronales artificiales.

    • 1997 – Deep Blue de IBM derrota al campeón mundial de ajedrez Kasparov, siendo la primera vez que una computadora vence a un campeón mundial en un juego complejo.
    • 2002 – iRobot presenta Roomba, el primer robot aspirador doméstico producido en serie con un sistema de navegación impulsado por IA.
    • 2011 – Watson de IBM derrota a dos ex campeones de Jeopardy!.
    • 2012 – La startup de inteligencia artificial DeepMind desarrolla una red neuronal profunda que puede reconocer gatos en vídeos de YouTube.
    • 2014 – Facebook crea DeepFace, un sistema de reconocimiento facial que puede reconocer rostros con una precisión casi humana.

    (!!) DeepMind fue adquirida por Google en 2014 por 500 millones de dólares.

    • 2015 – AlphaGo, desarrollado por DeepMind, derrota al campeón mundial Lee Sedol en el juego de Go.
    • 2017 – AlphaZero de Google derrota a los mejores motores de ajedrez y shogi del mundo en una serie de partidas.
    • 2020 – OpenAI lanza GPT-3, lo que marca un avance significativo en el procesamiento del lenguaje natural.

    (!!) Procesamiento del lenguaje natural: enseña a las computadoras a comprender y utilizar el lenguaje humano mediante técnicas como el aprendizaje automático.

    • 2021 – AlphaFold2 de DeepMind resuelve el problema del plegamiento de proteínas, allanando el camino para nuevos descubrimientos de fármacos y avances médicos.
    • 2022 – Google despide al ingeniero Blake Lemoine por sus afirmaciones de que el modelo de lenguaje para aplicaciones de diálogo (LaMDA) de Google era sensible.
    • 2023 – Artistas presentaron una demanda colectiva contra Stability AI, DeviantArt y Mid-journey por usar Stable Diffusion para remezclar las obras protegidas por derechos de autor de millones de artistas.

    Gráfico: Open Tech / Genuine Impact

    Entradas relacionadas

    #ajedrez #AlphaFold2 #AlphaGo #AlphaZero #aprendizajeAutomático #artículo #artistas #aspirador #BlakeLemoine #ConferenciaDeDartmouth #copyright #DeanEdmonds #DeepBlue #DeepFace #DeepMind #DeviantArt #ELIZA #Facebook #gatos #GenuineImpact #Go #Google #GPS #GPT3 #gráfico #Hinton #IA #IBM #infografía #inteligenciaArtificial #iRobot #Jeopardy_ #Kasparov #LaMDA #LeeSedol #MarvinMinsky #McCarthy #McCullock #MidJourney #modelos #Newell #OpenTech #OpenAI #patrones #Perceptron #Pitts #plegamientoDeProteínas #predicciones #procesamientoDelLenguajeNatural #reconocimientoFacial #redesNeuronales #remezclar #robot #Rochester #Roomba #Rosenblatt #Rumelhart #Shannon #shogi #Simon #sistemaDeNavegación #SNAR #StabilityAI #StableDiffusion #testDeTuring #Turing #vídeos #Watson #Weizenbaum #Williams #YouTube

  22. DeepMind AI rivals the world’s smartest high schoolers at geometry - Enlarge / Demis Hassabis, CEO of DeepMind Technologies and developer of... - arstechnica.com/?p=1997186 #alphageometry #alphazero #deepmind #science #alphago #ai

  23. DeepMind AI rivals the world’s smartest high schoolers at geometry - Enlarge / Demis Hassabis, CEO of DeepMind Technologies and developer of... - arstechnica.com/?p=1997186 #alphageometry #alphazero #deepmind #science #alphago #ai

  24. #chess #siliconroad #alphazero A new DeepMind paper on chess-related topics (this time looking at solving fortresses and Penrose positions) using AlphaZero arxiv.org/pdf/2308.09175.pdf

  25. #chess #siliconroad #alphazero A new DeepMind paper on chess-related topics (this time looking at solving fortresses and Penrose positions) using AlphaZero arxiv.org/pdf/2308.09175.pdf

  26. We fear our advanced #AIs will find loopholes in our ethical principles and their prime directives, thus spiralling out of control.

    Is there a reason to fear this? Certainly it's something that almost invariably happens with smaller AIs and simpler tasks; a Tetris-playing agent will quickly learn to pause the game to avoid game over.

    These kinds of AIs will learn to perform the task through the path of the least resistance, go over the lowest fence.

    But with more complex #ML models this changes abruptly. Suddenly the easiest way to imitate human writing isn't to cheat and mock, it is to actually learn human thinking, logic, intuitive understanding of the physical world and so on. Because cheating has become prohibitively expensive. A #ChineseRoom holding all the possible combinations of questions and answers would be vastly larger than a function describing intelligent thought.

    And that is why we got true intelligence out of these language prediction models, just like we got the same in scaled-up #RL models previously.

    Once the task and the criteria of judgement of the task become complex enough, it becomes easier to not cheat, as cheating becomes computationally intractable.

    The same goes with our ethical frameworks. If we put ~20 #LLM chatbots to judge and rank different aspects of the RL-trained LLM performance, like coherence, factuality, morality, respect for truth, ...; we will get a model which learns to actually internalize these values instead of trying to somehow hide that it doesn't.

    Hiding and lying simply becomes too difficult, especially against a panel of machine judges who can see the internal thinking of the agent judged (as in chain-of-thought schemes).

    So, I think this is a risk, but it can be very easily managed.

    As we can now easily bootstrap RL training of these models with our existing models, it is almost trivial to achieve an unambigous #AGI in a relatively short time. I'm sure everyone is working on this already, so this isn't anything spectacularly new or innovative. It's just taking the same steps as previously taken from #AlphaGo to #AlphaZero and beyond, going so much above human level that it can't even be measured anymore.

  27. We fear our advanced #AIs will find loopholes in our ethical principles and their prime directives, thus spiralling out of control.

    Is there a reason to fear this? Certainly it's something that almost invariably happens with smaller AIs and simpler tasks; a Tetris-playing agent will quickly learn to pause the game to avoid game over.

    These kinds of AIs will learn to perform the task through the path of the least resistance, go over the lowest fence.

    But with more complex #ML models this changes abruptly. Suddenly the easiest way to imitate human writing isn't to cheat and mock, it is to actually learn human thinking, logic, intuitive understanding of the physical world and so on. Because cheating has become prohibitively expensive. A #ChineseRoom holding all the possible combinations of questions and answers would be vastly larger than a function describing intelligent thought.

    And that is why we got true intelligence out of these language prediction models, just like we got the same in scaled-up #RL models previously.

    Once the task and the criteria of judgement of the task become complex enough, it becomes easier to not cheat, as cheating becomes computationally intractable.

    The same goes with our ethical frameworks. If we put ~20 #LLM chatbots to judge and rank different aspects of the RL-trained LLM performance, like coherence, factuality, morality, respect for truth, ...; we will get a model which learns to actually internalize these values instead of trying to somehow hide that it doesn't.

    Hiding and lying simply becomes too difficult, especially against a panel of machine judges who can see the internal thinking of the agent judged (as in chain-of-thought schemes).

    So, I think this is a risk, but it can be very easily managed.

    As we can now easily bootstrap RL training of these models with our existing models, it is almost trivial to achieve an unambigous #AGI in a relatively short time. I'm sure everyone is working on this already, so this isn't anything spectacularly new or innovative. It's just taking the same steps as previously taken from #AlphaGo to #AlphaZero and beyond, going so much above human level that it can't even be measured anymore.

  28. If #AI start writing undetectable malware (darkreading.com/attacks-breach) then what's the long term solution? Right now we probably don't have adequate defenses against many different vectors, because we haven't imagined them yet. But advanced AI might, just like how #AlphaZero discovered new strategies for Chess and Go.

    Maybe it's a time for #OpenSource to shine, coupled with firewall AI that reads and analyzes the code before running.

    #security #programming

  29. If #AI start writing undetectable malware (darkreading.com/attacks-breach) then what's the long term solution? Right now we probably don't have adequate defenses against many different vectors, because we haven't imagined them yet. But advanced AI might, just like how #AlphaZero discovered new strategies for Chess and Go.

    Maybe it's a time for #OpenSource to shine, coupled with firewall AI that reads and analyzes the code before running.

    #security #programming

  30. Holy Molly, it turns out that DeepMind have quietly open-sourced #mctx, the Monte Carlo search engine behind their #AlphaGo, #AlphaZero, and #MuZero #Go engines

    #go #baduk #weiqi

    github.com/deepmind/mctx

  31. Holy Molly, it turns out that DeepMind have quietly open-sourced #mctx, the Monte Carlo search engine behind their #AlphaGo, #AlphaZero, and #MuZero #Go engines

    #go #baduk #weiqi

    github.com/deepmind/mctx

  32. DeepMind has done really well at training game-playing neural network systems such as #AlphaZero. Who is exploring combining the ideas of AlphaZero's self-play training with LLM networks to help #chatgpt avoid hallucinations?

  33. DeepMind has done really well at training game-playing neural network systems such as #AlphaZero. Who is exploring combining the ideas of AlphaZero's self-play training with LLM networks to help #chatgpt avoid hallucinations?

  34. One thing which hinders #LLM #chatbot performance is that they are trained to imitate humans. Hence they tend to be bad at similar things humans are bad at.

    Reinforcement Learning with Human Feedback (#RLHF) improves this slightly by making the system compete against itself, where the performances are ranked by humans. After all, humans are better at ranking outputs than producing example outputs.

    It is possible to scale that up and maybe even improve over that slightly by utilizing the trained critic network to rank the performances "as if they had been ranked by humans", but that critic then imitates humans again with similar issues. RLHF typically uses a critic anyhow.

    Instead of, or in addition to that, we can make other games for these chatbots and train them with self-competition much like #AlphaZero/#MuZero. We can formulate all kinds of complex games and procedural challenges for the LLM, and make it compete against itself in such tasks which are easy to rank or evaluate algorithmically.

    Even playing #chess against itself would probably improve its skills not only in chess, but in a generalizable fashion to other tasks which require #planning.

    RLHF is used for things where humans are needed as "referees". However, this scales badly and is limited by human capability.

    Many games such as chess, or math problems, or playing #ATARI games in text, or controlling power plants by text can be "refereed" and scored automatically by machines.

    These types of problems will allow LLM chatbots to achieve superhuman capabilities not only in those tasks, but generally in other things they do as well, because the acquired skills are typically generally useful.

    The only requirement in addition to be scoreable automatically is that the games and challenges need to be presentable and played through text.

    #ChatGPT #LargeLanguageModels #DeepLearning #ReinforcementLearning

  35. One thing which hinders #LLM #chatbot performance is that they are trained to imitate humans. Hence they tend to be bad at similar things humans are bad at.

    Reinforcement Learning with Human Feedback (#RLHF) improves this slightly by making the system compete against itself, where the performances are ranked by humans. After all, humans are better at ranking outputs than producing example outputs.

    It is possible to scale that up and maybe even improve over that slightly by utilizing the trained critic network to rank the performances "as if they had been ranked by humans", but that critic then imitates humans again with similar issues. RLHF typically uses a critic anyhow.

    Instead of, or in addition to that, we can make other games for these chatbots and train them with self-competition much like #AlphaZero/#MuZero. We can formulate all kinds of complex games and procedural challenges for the LLM, and make it compete against itself in such tasks which are easy to rank or evaluate algorithmically.

    Even playing #chess against itself would probably improve its skills not only in chess, but in a generalizable fashion to other tasks which require #planning.

    RLHF is used for things where humans are needed as "referees". However, this scales badly and is limited by human capability.

    Many games such as chess, or math problems, or playing #ATARI games in text, or controlling power plants by text can be "refereed" and scored automatically by machines.

    These types of problems will allow LLM chatbots to achieve superhuman capabilities not only in those tasks, but generally in other things they do as well, because the acquired skills are typically generally useful.

    The only requirement in addition to be scoreable automatically is that the games and challenges need to be presentable and played through text.

    #ChatGPT #LargeLanguageModels #DeepLearning #ReinforcementLearning

  36. Google's #DeepMind has just open-sourced #AlphaGo. The Go reinforcement learning AI that beat the best Go player in the world a couple of years ago. Before this, researchers could access the algorithm only by sending a request to DeepMind, I suppose, because the researchers of DeepMind have published a paper in #Nature about AlphaGo and #AlphaZero. That's why there is a spinoff of AlaphaZero called Leela that was developed by someone who is totally unrelated to #Google.
    news.ycombinator.com/item?id=3

  37. DeepMind has open-sourced the core of #AlphaGo and #AlphaZero. It’s a library called mccx and provides JAX-native implementation of Monte Carlo Tree Search

    github.com/deepmind/mctx

  38. DeepMind has open-sourced the core of #AlphaGo and #AlphaZero. It’s a library called mccx and provides JAX-native implementation of Monte Carlo Tree Search

    github.com/deepmind/mctx

  39. I wrote this four years ago. Today, we're much closer to the reality I described than I thought we'd be…

    #AlphaZero, #machine #learning, and the #future of #work is.gd/10zWff

  40. I feel like I just beat #AlphaZero at chess and it kicked over the chessboard.

    Not only did #chatgpt have to apologize, it tried gaslighting its poor judgement about history and then knocked over the table....

  41. I feel like I just beat #AlphaZero at chess and it kicked over the chessboard.

    Not only did #chatgpt have to apologize, it tried gaslighting its poor judgement about history and then knocked over the table....

  42. There have been a few key moments in the recent history of AI: the Hinton et al's "vision" neural network; Watson's Jeopardy! win; AlphaZero's chess prowess. Any others you would add? Is OpenAI's chatbot in the same category? #ai #chatbot #alphazero #watson

  43. Mit Deep Reinforcement Learning hat DeepMind einen Algorithmus entdeckt, auf den kein Mensch kam. Er soll die Matrixmultiplikation signifikant beschleunigen.
    AlphaTensor: KI-System beschleunigt Matrixmultiplikation mit neuem Algorithmus