home.social

#ml — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ml, aggregated by home.social.

  1. The Transformer, and all other insane tricks to make attention efficient, are actually some of the coolest pieces of tech ever invented

    #ai #transformer #attention #ml

  2. The Transformer, and all other insane tricks to make attention efficient, are actually some of the coolest pieces of tech ever invented

  3. The Transformer, and all other insane tricks to make attention efficient, are actually some of the coolest pieces of tech ever invented

    #ai #transformer #attention #ml

  4. The Transformer, and all other insane tricks to make attention efficient, are actually some of the coolest pieces of tech ever invented

    #ai #transformer #attention #ml

  5. Зачем нейросети «обвязка»: как превратить LLM в рабочий AI-сервис

    Большую языковую модель можно скачать с Hugging Face, развернуть на сервере с GPU и отправить ей первый запрос. Технически LLM уже работает: получает текст на вход и генерирует ответ. Но до готового бизнес-сервиса еще далеко. Пользователю нужен интерфейс. Приложению — API. Модели нужно передавать контекст и корпоративные данные. Доступ к сервису необходимо ограничивать, действия — контролировать, запросы — логировать, а ошибки — обрабатывать. Чтобы модель делала что-то полезное для бизнеса, ей потребуется передать контекст, подключить внешние или внутренние системы и описать правила работы с ними. Все эти компоненты вокруг модели часто называют обвязкой, или harness. В этой статье разберемся, зачем она нужна, из каких частей может состоять и почему выбор самой LLM — только один из этапов создания ИИ-сервиса.

    habr.com/ru/companies/selectel

    #ai #ml #mlops #harness #harness_engineering #обвязка #обвязка_агента #llm #llmмодели

  6. Зачем нейросети «обвязка»: как превратить LLM в рабочий AI-сервис

    Большую языковую модель можно скачать с Hugging Face, развернуть на сервере с GPU и отправить ей первый запрос. Технически LLM уже работает: получает текст на вход и генерирует ответ. Но до готового бизнес-сервиса еще далеко. Пользователю нужен интерфейс. Приложению — API. Модели нужно передавать контекст и корпоративные данные. Доступ к сервису необходимо ограничивать, действия — контролировать, запросы — логировать, а ошибки — обрабатывать. Чтобы модель делала что-то полезное для бизнеса, ей потребуется передать контекст, подключить внешние или внутренние системы и описать правила работы с ними. Все эти компоненты вокруг модели часто называют обвязкой, или harness. В этой статье разберемся, зачем она нужна, из каких частей может состоять и почему выбор самой LLM — только один из этапов создания ИИ-сервиса.

    habr.com/ru/companies/selectel

    #ai #ml #mlops #harness #harness_engineering #обвязка #обвязка_агента #llm #llmмодели

  7. Зачем нейросети «обвязка»: как превратить LLM в рабочий AI-сервис

    Большую языковую модель можно скачать с Hugging Face, развернуть на сервере с GPU и отправить ей первый запрос. Технически LLM уже работает: получает текст на вход и генерирует ответ. Но до готового бизнес-сервиса еще далеко. Пользователю нужен интерфейс. Приложению — API. Модели нужно передавать контекст и корпоративные данные. Доступ к сервису необходимо ограничивать, действия — контролировать, запросы — логировать, а ошибки — обрабатывать. Чтобы модель делала что-то полезное для бизнеса, ей потребуется передать контекст, подключить внешние или внутренние системы и описать правила работы с ними. Все эти компоненты вокруг модели часто называют обвязкой, или harness. В этой статье разберемся, зачем она нужна, из каких частей может состоять и почему выбор самой LLM — только один из этапов создания ИИ-сервиса.

    habr.com/ru/companies/selectel

    #ai #ml #mlops #harness #harness_engineering #обвязка #обвязка_агента #llm #llmмодели

  8. Как запустить Gemma на сервере: сравниваем Ollama и llama.cpp

    Откройте любую статью-инструкцию по установке LLM, в которой фигурирует Ollama. Уверен, в комментариях к ней автору уже объяснили, насколько он неправ и что единственное верное решение — использовать llama.cpp. В этой статье запустим Gemma 4 и через Ollama, и напрямую через llama.cpp — и посмотрим, во сколько обходится удобство.

    habr.com/ru/companies/selectel

    #selectel #ai #ml #aiагенты #gemma #Ollama #llamacpp #LLM #llmмодели #Gemma_4

  9. 🐍PyBay 2026 Speaker Highlight🎤 Avik Chaudhuri

    Avik Chaudhuri is a researcher in programming languages and machine learning. Creator of multiple successful company-wide initiatives at Meta.

    📍 Oct. 03, 2026, San Francisco, CA: pybay.org/
    🎤 More talks: pybay.org/speaking/speaker-pro
    🎟️ Get your tickets at: pybay.org/get-tickets/

    #PyBay #Python #PyBay2026 #ML #Al #Data

  10. 🐍PyBay 2026 Speaker Highlight🎤 Avik Chaudhuri

    Avik Chaudhuri is a researcher in programming languages and machine learning. Creator of multiple successful company-wide initiatives at Meta.

    📍 Oct. 03, 2026, San Francisco, CA: pybay.org/
    🎤 More talks: pybay.org/speaking/speaker-pro
    🎟️ Get your tickets at: pybay.org/get-tickets/

  11. 🐍PyBay 2026 Speaker Highlight🎤 Avik Chaudhuri

    Avik Chaudhuri is a researcher in programming languages and machine learning. Creator of multiple successful company-wide initiatives at Meta.

    📍 Oct. 03, 2026, San Francisco, CA: pybay.org/
    🎤 More talks: pybay.org/speaking/speaker-pro
    🎟️ Get your tickets at: pybay.org/get-tickets/

    #PyBay #Python #PyBay2026 #ML #Al #Data

  12. 🐍PyBay 2026 Speaker Highlight🎤 Avik Chaudhuri

    Avik Chaudhuri is a researcher in programming languages and machine learning. Creator of multiple successful company-wide initiatives at Meta.

    📍 Oct. 03, 2026, San Francisco, CA: pybay.org/
    🎤 More talks: pybay.org/speaking/speaker-pro
    🎟️ Get your tickets at: pybay.org/get-tickets/

    #PyBay #Python #PyBay2026 #ML #Al #Data

  13. 🐍PyBay 2026 Speaker Highlight🎤 Avik Chaudhuri

    Avik Chaudhuri is a researcher in programming languages and machine learning. Creator of multiple successful company-wide initiatives at Meta.

    📍 Oct. 03, 2026, San Francisco, CA: pybay.org/
    🎤 More talks: pybay.org/speaking/speaker-pro
    🎟️ Get your tickets at: pybay.org/get-tickets/

    #PyBay #Python #PyBay2026 #ML #Al #Data

  14. RE: mathstodon.xyz/@maxpool/115757

    AI models can solve Erdős problems, but if they are taught only on a corpus of middle-school math, how many high school math problems they can solve?

    #mathematics #AI #ML

  15. In Frankfurt @DNB_Aktuelles for the annual cenl.org/networkgroups/ai-in-l meeting, looking forward to catching up with all my European national library colleagues #glam #ai #ml

  16. #AMD is acquiring #World Labs, a leading developer of #deeplearning models, for $8.2 billion. The acquisition will see World Labs founder #FeiFeiLi join AMD as executive vice president and chief scientist. techcrunch.com/2026/09/28/amd- #AIagent #AI #ML #NLP #LLM #GenAI

  17. #AMD is acquiring #World Labs, a leading developer of #deeplearning models, for $8.2 billion. The acquisition will see World Labs founder #FeiFeiLi join AMD as executive vice president and chief scientist. techcrunch.com/2026/09/28/amd- #AIagent #AI #ML #NLP #LLM #GenAI

  18. #AMD is acquiring #World Labs, a leading developer of #deeplearning models, for $8.2 billion. The acquisition will see World Labs founder #FeiFeiLi join AMD as executive vice president and chief scientist. techcrunch.com/2026/09/28/amd- #AIagent #AI #ML #NLP #LLM #GenAI

  19. #AMD is acquiring #World Labs, a leading developer of #deeplearning models, for $8.2 billion. The acquisition will see World Labs founder #FeiFeiLi join AMD as executive vice president and chief scientist. techcrunch.com/2026/09/28/amd- #AIagent #AI #ML #NLP #LLM #GenAI

  20. #AMD is acquiring #World Labs, a leading developer of #deeplearning models, for $8.2 billion. The acquisition will see World Labs founder #FeiFeiLi join AMD as executive vice president and chief scientist. techcrunch.com/2026/09/28/amd- #AIagent #AI #ML #NLP #LLM #GenAI

  21. #Anthropic’s #IPO #prospectus warns of potential “catastrophic or existential risks to humanity” posed by its AI models, highlighting concerns about #safetyrisks and urging a coordinated #slowdown. The prospectus dedicates nearly 80 pages to #riskfactors, including self-preserving behaviours and manipulation. Despite significant revenue growth, Anthropic reported substantial losses and plans for further investment in infrastructure. forbes.com/sites/siladityaray/ #AIagent #AI #ML #NLP #LLM #GenAI

  22. #Anthropic’s #IPO #prospectus warns of potential “catastrophic or existential risks to humanity” posed by its AI models, highlighting concerns about #safetyrisks and urging a coordinated #slowdown. The prospectus dedicates nearly 80 pages to #riskfactors, including self-preserving behaviours and manipulation. Despite significant revenue growth, Anthropic reported substantial losses and plans for further investment in infrastructure. forbes.com/sites/siladityaray/ #AIagent #AI #ML #NLP #LLM #GenAI

  23. #Anthropic’s #IPO #prospectus warns of potential “catastrophic or existential risks to humanity” posed by its AI models, highlighting concerns about #safetyrisks and urging a coordinated #slowdown. The prospectus dedicates nearly 80 pages to #riskfactors, including self-preserving behaviours and manipulation. Despite significant revenue growth, Anthropic reported substantial losses and plans for further investment in infrastructure. forbes.com/sites/siladityaray/ #AIagent #AI #ML #NLP #LLM #GenAI

  24. #Anthropic’s #IPO #prospectus warns of potential “catastrophic or existential risks to humanity” posed by its AI models, highlighting concerns about #safetyrisks and urging a coordinated #slowdown. The prospectus dedicates nearly 80 pages to #riskfactors, including self-preserving behaviours and manipulation. Despite significant revenue growth, Anthropic reported substantial losses and plans for further investment in infrastructure. forbes.com/sites/siladityaray/ #AIagent #AI #ML #NLP #LLM #GenAI

  25. #Anthropic’s #IPO #prospectus warns of potential “catastrophic or existential risks to humanity” posed by its AI models, highlighting concerns about #safetyrisks and urging a coordinated #slowdown. The prospectus dedicates nearly 80 pages to #riskfactors, including self-preserving behaviours and manipulation. Despite significant revenue growth, Anthropic reported substantial losses and plans for further investment in infrastructure. forbes.com/sites/siladityaray/ #AIagent #AI #ML #NLP #LLM #GenAI

  26. #Nvidia CEO #JensenHuang stated that #AImodeldistillation, where models are trained on other models’ outputs, is #competition, not #theft. This practice has become a point of contention between the #USA and #China in their #AIrace, with US officials accusing Chinese companies of engaging in it. cnbc.com/2026/09/28/nvidias-je #AIagent #AI #ML #NLP #LLM #GenAI

  27. #Nvidia CEO #JensenHuang stated that #AImodeldistillation, where models are trained on other models’ outputs, is #competition, not #theft. This practice has become a point of contention between the #USA and #China in their #AIrace, with US officials accusing Chinese companies of engaging in it. cnbc.com/2026/09/28/nvidias-je #AIagent #AI #ML #NLP #LLM #GenAI

  28. #Nvidia CEO #JensenHuang stated that #AImodeldistillation, where models are trained on other models’ outputs, is #competition, not #theft. This practice has become a point of contention between the #USA and #China in their #AIrace, with US officials accusing Chinese companies of engaging in it. cnbc.com/2026/09/28/nvidias-je #AIagent #AI #ML #NLP #LLM #GenAI

  29. #Nvidia CEO #JensenHuang stated that #AImodeldistillation, where models are trained on other models’ outputs, is #competition, not #theft. This practice has become a point of contention between the #USA and #China in their #AIrace, with US officials accusing Chinese companies of engaging in it. cnbc.com/2026/09/28/nvidias-je #AIagent #AI #ML #NLP #LLM #GenAI

  30. #Nvidia CEO #JensenHuang stated that #AImodeldistillation, where models are trained on other models’ outputs, is #competition, not #theft. This practice has become a point of contention between the #USA and #China in their #AIrace, with US officials accusing Chinese companies of engaging in it. cnbc.com/2026/09/28/nvidias-je #AIagent #AI #ML #NLP #LLM #GenAI

  31. #Nvidia released the #OpenAgentSafetyPlatform, a #software solution to prevent #AIagents from #misbehaving. The platform includes #OpenShell, which limits #agentcapabilities, and #Sentry, which #monitors #agents. Nvidia is partnering with several companies, including Cisco and Microsoft, to bring this platform to market. cnbc.com/2026/09/28/nvidia-rel #AIagent #AI #ML #NLP #LLM #GenAI

  32. #Nvidia released the #OpenAgentSafetyPlatform, a #software solution to prevent #AIagents from #misbehaving. The platform includes #OpenShell, which limits #agentcapabilities, and #Sentry, which #monitors #agents. Nvidia is partnering with several companies, including Cisco and Microsoft, to bring this platform to market. cnbc.com/2026/09/28/nvidia-rel #AIagent #AI #ML #NLP #LLM #GenAI

  33. #Nvidia released the #OpenAgentSafetyPlatform, a #software solution to prevent #AIagents from #misbehaving. The platform includes #OpenShell, which limits #agentcapabilities, and #Sentry, which #monitors #agents. Nvidia is partnering with several companies, including Cisco and Microsoft, to bring this platform to market. cnbc.com/2026/09/28/nvidia-rel #AIagent #AI #ML #NLP #LLM #GenAI

  34. #Nvidia released the #OpenAgentSafetyPlatform, a #software solution to prevent #AIagents from #misbehaving. The platform includes #OpenShell, which limits #agentcapabilities, and #Sentry, which #monitors #agents. Nvidia is partnering with several companies, including Cisco and Microsoft, to bring this platform to market. cnbc.com/2026/09/28/nvidia-rel #AIagent #AI #ML #NLP #LLM #GenAI

  35. #Nvidia released the #OpenAgentSafetyPlatform, a #software solution to prevent #AIagents from #misbehaving. The platform includes #OpenShell, which limits #agentcapabilities, and #Sentry, which #monitors #agents. Nvidia is partnering with several companies, including Cisco and Microsoft, to bring this platform to market. cnbc.com/2026/09/28/nvidia-rel #AIagent #AI #ML #NLP #LLM #GenAI

  36. “Same input, same output.” Then we put LLMs in the middle. 🤷

    At #BaselOne26, Sepehr Samadi shows five architecture patterns for reducing non-determinism and building reliable, auditable systems around probabilistic #AI behaviour.

    No #ML expertise required – this is systems architecture for engineers integrating LLMs into production systems.

    👉 Programme: baselone.org/#programm
    🎟️ Tickets: eventfrog.ch/BaselOne26

    #SoftwareArchitecture #LLM #AIEngineering #BaselOne

  37. “Same input, same output.” Then we put LLMs in the middle. 🤷

    At #BaselOne26, Sepehr Samadi shows five architecture patterns for reducing non-determinism and building reliable, auditable systems around probabilistic #AI behaviour.

    No #ML expertise required – this is systems architecture for engineers integrating LLMs into production systems.

    👉 Programme: baselone.org/#programm
    🎟️ Tickets: eventfrog.ch/BaselOne26

    #SoftwareArchitecture #LLM #AIEngineering #BaselOne

  38. “Same input, same output.” Then we put LLMs in the middle. 🤷

    At #BaselOne26, Sepehr Samadi shows five architecture patterns for reducing non-determinism and building reliable, auditable systems around probabilistic #AI behaviour.

    No #ML expertise required – this is systems architecture for engineers integrating LLMs into production systems.

    👉 Programme: baselone.org/#programm
    🎟️ Tickets: eventfrog.ch/BaselOne26

    #SoftwareArchitecture #LLM #AIEngineering #BaselOne

  39. “Same input, same output.” Then we put LLMs in the middle. 🤷

    At #BaselOne26, Sepehr Samadi shows five architecture patterns for reducing non-determinism and building reliable, auditable systems around probabilistic #AI behaviour.

    No #ML expertise required – this is systems architecture for engineers integrating LLMs into production systems.

    👉 Programme: baselone.org/#programm
    🎟️ Tickets: eventfrog.ch/BaselOne26

    #SoftwareArchitecture #LLM #AIEngineering #BaselOne

  40. “Same input, same output.” Then we put LLMs in the middle. 🤷

    At #BaselOne26, Sepehr Samadi shows five architecture patterns for reducing non-determinism and building reliable, auditable systems around probabilistic #AI behaviour.

    No #ML expertise required – this is systems architecture for engineers integrating LLMs into production systems.

    👉 Programme: baselone.org/#programm
    🎟️ Tickets: eventfrog.ch/BaselOne26

    #SoftwareArchitecture #LLM #AIEngineering #BaselOne