home.social

#eleutherai — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #eleutherai, aggregated by home.social.

fetched live
  1. 🔍 Nghiên cứu phân tích 23+ mô hình từ 7 phòng thí nghiệm cho thấy đặc tính “thermodynamic” của mô hình phụ thuộc nhiều hơn vào nhà phát triển hơn là số tham số. Các mô hình EleutherAI (Pythia, GPT‑NeoX) có xu hướng giảm tín hiệu (G<1), trong khi Meta/OpenAI (LLaMA, OPT, GPT‑2) mở rộng (G>1). Fine‑tuning chỉ thay đổi độ lớn, hiếm khi đảo ngược dấu, nên việc chọn base model quan trọng. #AI #NLP #LLM #MachineLearning #Mô_hình #DeepLearning #Research #EleutherAI #Meta #LLaMA #Finetuning

    https://ww

  2. TechCrunch: EleutherAI releases massive AI training dataset of licensed and open domain text. “The dataset, called the Common Pile v0.1, took around two years to complete in collaboration with AI startups Poolside, Hugging Face, and others, along with several academic institutions. Weighing in at 8 terabytes in size, the Common Pile v0.1 was used to train two new AI models from EleutherAI, […]

    https://rbfirehose.com/2025/06/07/techcrunch-eleutherai-releases-massive-ai-training-dataset-of-licensed-and-open-domain-text/

  3. TechCrunch: EleutherAI releases massive AI training dataset of licensed and open domain text. “The dataset, called the Common Pile v0.1, took around two years to complete in collaboration with AI startups Poolside, Hugging Face, and others, along with several academic institutions. Weighing in at 8 terabytes in size, the Common Pile v0.1 was used to train two new AI models from EleutherAI, […]

    https://rbfirehose.com/2025/06/07/techcrunch-eleutherai-releases-massive-ai-training-dataset-of-licensed-and-open-domain-text/

  4. TechCrunch: EleutherAI releases massive AI training dataset of licensed and open domain text. “The dataset, called the Common Pile v0.1, took around two years to complete in collaboration with AI startups Poolside, Hugging Face, and others, along with several academic institutions. Weighing in at 8 terabytes in size, the Common Pile v0.1 was used to train two new AI models from EleutherAI, […]

    https://rbfirehose.com/2025/06/07/techcrunch-eleutherai-releases-massive-ai-training-dataset-of-licensed-and-open-domain-text/

  5. TechCrunch: EleutherAI releases massive AI training dataset of licensed and open domain text. “The dataset, called the Common Pile v0.1, took around two years to complete in collaboration with AI startups Poolside, Hugging Face, and others, along with several academic institutions. Weighing in at 8 terabytes in size, the Common Pile v0.1 was used to train two new AI models from EleutherAI, […]

    https://rbfirehose.com/2025/06/07/techcrunch-eleutherai-releases-massive-ai-training-dataset-of-licensed-and-open-domain-text/

  6. TechCrunch: EleutherAI releases massive AI training dataset of licensed and open domain text. “The dataset, called the Common Pile v0.1, took around two years to complete in collaboration with AI startups Poolside, Hugging Face, and others, along with several academic institutions. Weighing in at 8 terabytes in size, the Common Pile v0.1 was used to train two new AI models from EleutherAI, […]

    https://rbfirehose.com/2025/06/07/techcrunch-eleutherai-releases-massive-ai-training-dataset-of-licensed-and-open-domain-text/

  7. EleutherAI is a grassroots non-profit AI research group, formed in July 2020 by Connor Leahy, Sid Black, and Leo Gao. Known for creating open-source models like GPT-Neo, GPT-J, and GPT-NeoX, their Pile dataset is widely used for training large language models. In early 2023, they incorporated as the EleutherAI Institute. #AI #OpenSource #EleutherAI #MachineLearning #GPT
    eleuther.ai

  8. EleutherAI is a grassroots non-profit AI research group, formed in July 2020 by Connor Leahy, Sid Black, and Leo Gao. Known for creating open-source models like GPT-Neo, GPT-J, and GPT-NeoX, their Pile dataset is widely used for training large language models. In early 2023, they incorporated as the EleutherAI Institute. #AI #OpenSource #EleutherAI #MachineLearning #GPT
    eleuther.ai

  9. EleutherAI is a grassroots non-profit AI research group, formed in July 2020 by Connor Leahy, Sid Black, and Leo Gao. Known for creating open-source models like GPT-Neo, GPT-J, and GPT-NeoX, their Pile dataset is widely used for training large language models. In early 2023, they incorporated as the EleutherAI Institute. #AI #OpenSource #EleutherAI #MachineLearning #GPT
    eleuther.ai

  10. EleutherAI is a grassroots non-profit AI research group, formed in July 2020 by Connor Leahy, Sid Black, and Leo Gao. Known for creating open-source models like GPT-Neo, GPT-J, and GPT-NeoX, their Pile dataset is widely used for training large language models. In early 2023, they incorporated as the EleutherAI Institute. #AI #OpenSource #EleutherAI #MachineLearning #GPT
    eleuther.ai

  11. EleutherAI is a grassroots non-profit AI research group, formed in July 2020 by Connor Leahy, Sid Black, and Leo Gao. Known for creating open-source models like GPT-Neo, GPT-J, and GPT-NeoX, their Pile dataset is widely used for training large language models. In early 2023, they incorporated as the EleutherAI Institute. #AI #OpenSource #EleutherAI #MachineLearning #GPT
    eleuther.ai

  12. "Recently it was revealed that an AI research lab called #EleutherAI had harvested subtitles from YouTube videos without the creators' consent. This data was then combined with data from Wikipedia, the U.K. Parliament and Enron Staff emails and added to a dataset called “the Pile.”
    (Tom's Guide 7/22/2024)

  13. "Recently it was revealed that an AI research lab called #EleutherAI had harvested subtitles from YouTube videos without the creators' consent. This data was then combined with data from Wikipedia, the U.K. Parliament and Enron Staff emails and added to a dataset called “the Pile.”
    (Tom's Guide 7/22/2024)

  14. YouTube creators surprised to find Apple and others trained AI on their videos - Enlarge / YouTuber Marques Brownlee discusses iOS 18 in a new video. Th... - arstechnica.com/?p=2037316 #largelanguagemodels #eleutherai #anthropic #thepile #youtube #google #apple #tech #ai

  15. YouTube creators surprised to find Apple and others trained AI on their videos - Enlarge / YouTuber Marques Brownlee discusses iOS 18 in a new video. Th... - arstechnica.com/?p=2037316 #largelanguagemodels #eleutherai #anthropic #thepile #youtube #google #apple #tech #ai

  16. YouTube creators surprised to find Apple and others trained AI on their videos - Enlarge / YouTuber Marques Brownlee discusses iOS 18 in a new video. Th... - arstechnica.com/?p=2037316 #largelanguagemodels #eleutherai #anthropic #thepile #youtube #google #apple #tech #ai

  17. YouTube creators surprised to find Apple and others trained AI on their videos - Enlarge / YouTuber Marques Brownlee discusses iOS 18 in a new video. Th... - arstechnica.com/?p=2037316 #largelanguagemodels #eleutherai #anthropic #thepile #youtube #google #apple #tech #ai

  18. YouTube creators surprised to find Apple and others trained AI on their videos - Enlarge / YouTuber Marques Brownlee discusses iOS 18 in a new video. Th... - arstechnica.com/?p=2037316 #largelanguagemodels #eleutherai #anthropic #thepile #youtube #google #apple #tech #ai

  19. How the Foundation Model Transparency Index Distorts Transparency | EleutherAI Blog blog.eleuther.ai/fmti-critique

    I saw the Foundation Model Transparency Index paper come out recently and was surprised that OpenAI scored as high as they did. This Eleuther AI post breaks down how the Foundation Model Transparency index gets it all wrong, and is not really measuring transparency at all.

    #fmti
    #foundationmodeltransparencyindex
    #opensource
    #LLM
    #eleutherai

  20. How the Foundation Model Transparency Index Distorts Transparency | EleutherAI Blog blog.eleuther.ai/fmti-critique

    I saw the Foundation Model Transparency Index paper come out recently and was surprised that OpenAI scored as high as they did. This Eleuther AI post breaks down how the Foundation Model Transparency index gets it all wrong, and is not really measuring transparency at all.

    #fmti
    #foundationmodeltransparencyindex
    #opensource
    #LLM
    #eleutherai

  21. How the Foundation Model Transparency Index Distorts Transparency | EleutherAI Blog blog.eleuther.ai/fmti-critique

    I saw the Foundation Model Transparency Index paper come out recently and was surprised that OpenAI scored as high as they did. This Eleuther AI post breaks down how the Foundation Model Transparency index gets it all wrong, and is not really measuring transparency at all.

    #fmti
    #foundationmodeltransparencyindex
    #opensource
    #LLM
    #eleutherai

  22. How the Foundation Model Transparency Index Distorts Transparency | EleutherAI Blog blog.eleuther.ai/fmti-critique

    I saw the Foundation Model Transparency Index paper come out recently and was surprised that OpenAI scored as high as they did. This Eleuther AI post breaks down how the Foundation Model Transparency index gets it all wrong, and is not really measuring transparency at all.

    #fmti
    #foundationmodeltransparencyindex
    #opensource
    #LLM
    #eleutherai

  23. CW: Long thread/40

    Some "open AI" is much more open than the industry dominating offerings. There's #EleutherAI, a donor-supported nonprofit whose model comes with documentation and code, licensed #Apache2. There are also some smaller academic offerings: #Vicuna (UCSD/CMU/Berkeley); #Koala (Berkeley) and #Alpaca (Stanford).

    These are indeed more open (though Alpaca - which ran on a laptop - had to be withdrawn because it "hallucinated" so profusely).

    40/

  24. CW: Long thread/40

    Some "open AI" is much more open than the industry dominating offerings. There's #EleutherAI, a donor-supported nonprofit whose model comes with documentation and code, licensed #Apache2. There are also some smaller academic offerings: #Vicuna (UCSD/CMU/Berkeley); #Koala (Berkeley) and #Alpaca (Stanford).

    These are indeed more open (though Alpaca - which ran on a laptop - had to be withdrawn because it "hallucinated" so profusely).

    40/

  25. CW: Long thread/40

    Some "open AI" is much more open than the industry dominating offerings. There's #EleutherAI, a donor-supported nonprofit whose model comes with documentation and code, licensed #Apache2. There are also some smaller academic offerings: #Vicuna (UCSD/CMU/Berkeley); #Koala (Berkeley) and #Alpaca (Stanford).

    These are indeed more open (though Alpaca - which ran on a laptop - had to be withdrawn because it "hallucinated" so profusely).

    40/

  26. CW: Long thread/40

    Some "open AI" is much more open than the industry dominating offerings. There's #EleutherAI, a donor-supported nonprofit whose model comes with documentation and code, licensed #Apache2. There are also some smaller academic offerings: #Vicuna (UCSD/CMU/Berkeley); #Koala (Berkeley) and #Alpaca (Stanford).

    These are indeed more open (though Alpaca - which ran on a laptop - had to be withdrawn because it "hallucinated" so profusely).

    40/

  27. CW: Long thread/40

    Some "open AI" is much more open than the industry dominating offerings. There's #EleutherAI, a donor-supported nonprofit whose model comes with documentation and code, licensed #Apache2. There are also some smaller academic offerings: #Vicuna (UCSD/CMU/Berkeley); #Koala (Berkeley) and #Alpaca (Stanford).

    These are indeed more open (though Alpaca - which ran on a laptop - had to be withdrawn because it "hallucinated" so profusely).

    40/

  28. 🚀 New episode of The Changelog!

    This week we’re taking you to the hallway track of The #Linux Foundation’s #OSSummit North America 2023 in Vancouver, Canada 🇨🇦

    This episode features three conversations about #opensource #AI:

    1️⃣ Beyang Liu (Co-founder and CTO at #Sourcegraph)
    2️⃣ @dennyglee (Developer Advocate at #Databricks)
    3️⃣ Stella Biderman (Head of Research at #EleutherAI)

    🎧 changelog.fm/541

  29. 🚀 New episode of The Changelog!

    This week we’re taking you to the hallway track of The #Linux Foundation’s #OSSummit North America 2023 in Vancouver, Canada 🇨🇦

    This episode features three conversations about #opensource #AI:

    1️⃣ Beyang Liu (Co-founder and CTO at #Sourcegraph)
    2️⃣ @dennyglee (Developer Advocate at #Databricks)
    3️⃣ Stella Biderman (Head of Research at #EleutherAI)

    🎧 changelog.fm/541

  30. 🚀 New episode of The Changelog!

    This week we’re taking you to the hallway track of The #Linux Foundation’s #OSSummit North America 2023 in Vancouver, Canada 🇨🇦

    This episode features three conversations about #opensource #AI:

    1️⃣ Beyang Liu (Co-founder and CTO at #Sourcegraph)
    2️⃣ @dennyglee (Developer Advocate at #Databricks)
    3️⃣ Stella Biderman (Head of Research at #EleutherAI)

    🎧 changelog.fm/541

  31. 🚀 New episode of The Changelog!

    This week we’re taking you to the hallway track of The #Linux Foundation’s #OSSummit North America 2023 in Vancouver, Canada 🇨🇦

    This episode features three conversations about #opensource #AI:

    1️⃣ Beyang Liu (Co-founder and CTO at #Sourcegraph)
    2️⃣ @dennyglee (Developer Advocate at #Databricks)
    3️⃣ Stella Biderman (Head of Research at #EleutherAI)

    🎧 changelog.fm/541

  32. 🚀 New episode of The Changelog!

    This week we’re taking you to the hallway track of The #Linux Foundation’s #OSSummit North America 2023 in Vancouver, Canada 🇨🇦

    This episode features three conversations about #opensource #AI:

    1️⃣ Beyang Liu (Co-founder and CTO at #Sourcegraph)
    2️⃣ @dennyglee (Developer Advocate at #Databricks)
    3️⃣ Stella Biderman (Head of Research at #EleutherAI)

    🎧 changelog.fm/541

  33. Releasing 3B and 7B #RedPajama-#INCITE family of models including base, instruction-tuned & chat models — #TOGETHER

    "The biggest takeaway is the demonstration that performant #LLMs can be built quickly by the open-source community. This work builds on top of our 1.2 trillion token RedPajama dataset, EleutherAI’s #Pythia training code, #FlashAttention from #Stanford and #Together, the #HELM benchmarks from Stanford #CRFM and generous support from #MILA, #EleutherAI & #LAION for compute time on the #Summit #supercomputer within the INCITE program award 'Scalable Foundation Models for Transferable Generalist AI'. We believe these kind of open collaborations, at larger scales, will be behind the best #AI systems of the future. "

    together.xyz/blog/redpajama-mo

  34. Releasing 3B and 7B #RedPajama-#INCITE family of models including base, instruction-tuned & chat models — #TOGETHER

    "The biggest takeaway is the demonstration that performant #LLMs can be built quickly by the open-source community. This work builds on top of our 1.2 trillion token RedPajama dataset, EleutherAI’s #Pythia training code, #FlashAttention from #Stanford and #Together, the #HELM benchmarks from Stanford #CRFM and generous support from #MILA, #EleutherAI & #LAION for compute time on the #Summit #supercomputer within the INCITE program award 'Scalable Foundation Models for Transferable Generalist AI'. We believe these kind of open collaborations, at larger scales, will be behind the best #AI systems of the future. "

    together.xyz/blog/redpajama-mo

  35. Releasing 3B and 7B #RedPajama-#INCITE family of models including base, instruction-tuned & chat models — #TOGETHER

    "The biggest takeaway is the demonstration that performant #LLMs can be built quickly by the open-source community. This work builds on top of our 1.2 trillion token RedPajama dataset, EleutherAI’s #Pythia training code, #FlashAttention from #Stanford and #Together, the #HELM benchmarks from Stanford #CRFM and generous support from #MILA, #EleutherAI & #LAION for compute time on the #Summit #supercomputer within the INCITE program award 'Scalable Foundation Models for Transferable Generalist AI'. We believe these kind of open collaborations, at larger scales, will be behind the best #AI systems of the future. "

    together.xyz/blog/redpajama-mo

  36. Releasing 3B and 7B #RedPajama-#INCITE family of models including base, instruction-tuned & chat models — #TOGETHER

    "The biggest takeaway is the demonstration that performant #LLMs can be built quickly by the open-source community. This work builds on top of our 1.2 trillion token RedPajama dataset, EleutherAI’s #Pythia training code, #FlashAttention from #Stanford and #Together, the #HELM benchmarks from Stanford #CRFM and generous support from #MILA, #EleutherAI & #LAION for compute time on the #Summit #supercomputer within the INCITE program award 'Scalable Foundation Models for Transferable Generalist AI'. We believe these kind of open collaborations, at larger scales, will be behind the best #AI systems of the future. "

    together.xyz/blog/redpajama-mo

  37. Releasing 3B and 7B #RedPajama-#INCITE family of models including base, instruction-tuned & chat models — #TOGETHER

    "The biggest takeaway is the demonstration that performant #LLMs can be built quickly by the open-source community. This work builds on top of our 1.2 trillion token RedPajama dataset, EleutherAI’s #Pythia training code, #FlashAttention from #Stanford and #Together, the #HELM benchmarks from Stanford #CRFM and generous support from #MILA, #EleutherAI & #LAION for compute time on the #Summit #supercomputer within the INCITE program award 'Scalable Foundation Models for Transferable Generalist AI'. We believe these kind of open collaborations, at larger scales, will be behind the best #AI systems of the future. "

    together.xyz/blog/redpajama-mo

  38. “A really big deal”—Dolly is a free, open source, ChatGPT-style AI model - Enlarge (credit: Databricks)

    On Wednesday, Databricks released... - arstechnica.com/?p=1931693 #largelanguagemodels #machinelearning #textsynthesis #apachespark #databricks #eleutherai #finetuning #biz⁢ #pythia #dolly #llama #meta #ai

  39. “A really big deal”—Dolly is a free, open source, ChatGPT-style AI model - Enlarge (credit: Databricks)

    On Wednesday, Databricks released... - arstechnica.com/?p=1931693 #largelanguagemodels #machinelearning #textsynthesis #apachespark #databricks #eleutherai #finetuning #biz⁢ #pythia #dolly #llama #meta #ai

  40. “A really big deal”—Dolly is a free, open source, ChatGPT-style AI model - Enlarge (credit: Databricks)

    On Wednesday, Databricks released... - arstechnica.com/?p=1931693 #largelanguagemodels #machinelearning #textsynthesis #apachespark #databricks #eleutherai #finetuning #biz⁢ #pythia #dolly #llama #meta #ai

  41. “A really big deal”—Dolly is a free, open source, ChatGPT-style AI model - Enlarge (credit: Databricks)

    On Wednesday, Databricks released... - arstechnica.com/?p=1931693 #largelanguagemodels #machinelearning #textsynthesis #apachespark #databricks #eleutherai #finetuning #biz⁢ #pythia #dolly #llama #meta #ai

  42. There are some really good papers that have sought to make the best of the current situation, but #EleutherAI had the compute to do it the right way and so we did.

    arxiv.org/abs/2211.08411
    arxiv.org/abs/2202.07646
    arxiv.org/abs/2202.07206
    arxiv.org/abs/2207.14251

    We hope that this work will empower more people to work on questions in interpretability, especially the causal impact of training data on model behavior!

  43. There are some really good papers that have sought to make the best of the current situation, but #EleutherAI had the compute to do it the right way and so we did.

    arxiv.org/abs/2211.08411
    arxiv.org/abs/2202.07646
    arxiv.org/abs/2202.07206
    arxiv.org/abs/2207.14251

    We hope that this work will empower more people to work on questions in interpretability, especially the causal impact of training data on model behavior!