home.social

#stylometry — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #stylometry, aggregated by home.social.

fetched live
  1. Just to show I’ll give AI credit where credit is due…

    I had been pondering this from time to time over many years. AI gave a damn good answer and thera- err… debriefing session:

    As a freshman undergraduate psychology major at CMU in 1987, we were required to earn credit as subjects in graduate student experiments. I was in a brief, no more than an hour-long experiment in which I was typing on a terminal in a room on my own and told I was communicating with another student whom I never did meet and whose text lines to me came up on my screen. The topic we were told to discuss, I suspect was arbitrary, and had something to do with policies on stores selling alcohol to underage purchasers.

    In this experiment was I communicating with an early AI prototype?

    Also, all these years later, is it unusual that not having a final debriefing about the experiment bothers me slightly?

    Damn! Now I want to apologize to the researchers (again) for having been such a slow typist 😁.

    #AI #AI #ArtificialIntelligence #ChatGPT #chatbot #Claude #college #contextual #control #Copilot #courses #DeepLearning #DeepSeek #FOIA #ForensicLinguistics #game #Gemini #generative #GenerativeAI #Google #grammar #Grok #idiolect #language #largeLanguageModel #linguistics #LLM #machineLearning #NaturalLanguageProcessing #neuralNetwork #NLP #OpenAI #psycholinguistics #research #stylometry #syntax #text
  2. Just to show I’ll give AI credit where credit is due…

    I had been pondering this from time to time over many years. AI gave a damn good answer and thera- err… debriefing session:

    As a freshman undergraduate psychology major at CMU in 1987, we were required to earn credit as subjects in graduate student experiments. I was in a brief, no more than an hour-long experiment in which I was typing on a terminal in a room on my own and told I was communicating with another student whom I never did meet and whose text lines to me came up on my screen. The topic we were told to discuss, I suspect was arbitrary, and had something to do with policies on stores selling alcohol to underage purchasers.

    In this experiment was I communicating with an early AI prototype?

    Also, all these years later, is it unusual that not having a final debriefing about the experiment bothers me slightly?

    Damn! Now I want to apologize to the researchers (again) for having been such a slow typist 😁.

    #AI #AI #ArtificialIntelligence #ChatGPT #chatbot #Claude #college #contextual #control #Copilot #courses #DeepLearning #DeepSeek #FOIA #ForensicLinguistics #game #Gemini #generative #GenerativeAI #Google #grammar #Grok #idiolect #language #largeLanguageModel #linguistics #LLM #machineLearning #NaturalLanguageProcessing #neuralNetwork #NLP #OpenAI #psycholinguistics #research #stylometry #syntax #text
  3. Our lab presented five papers at the last European conference of the International Association for Forensic and Legal Linguistics in June in Montpellier. Here are some pictures as well as links to see the abstracts and slides of the talks.

    andreanini.com/2026/07/15/our-

    #forensiclinguistics #linguistics #authorship #authorshipanalysis #stylometry #digitalhumanities

  4. Our latest research (led by Sadie Barlow, with Dr Edoardo Manino) found LambdaG doesn’t need a calibration dataset for well-calibrated likelihood ratios, just a simple correction! This simplifies analysis and makes it the only authorship verification method without calibration. Paper on arXiv now.

    andreanini.com/2026/07/13/lamb

    #forensiclinguistics #forensicscience #nlp #authorshipanalysis #stylometry #digitalhumanities

  5. #Romanistiktag #Konstanz heute morgen meine Lieblingsthemen in unserer Sektion: romanistiktag.de/xxxix-romanis : unsere Keynote von la famosa Laura Hernández Lopez: zur morphosyntaktischen, stilistisch und stilometrischen Analyse der Poesie von Autorinnen des Siglo de Oro. #stylometry #cls #dh

  6. #Romanistiktag #Konstanz heute morgen meine Lieblingsthemen in unserer Sektion: romanistiktag.de/xxxix-romanis : unsere Keynote von la famosa Laura Hernández Lopez: zur morphosyntaktischen, stilistisch und stilometrischen Analyse der Poesie von Autorinnen des Siglo de Oro. #stylometry #cls #dh

  7. My own talk, which is coming up / just happened, is on #Multilingual #Stylometry.

    Based on our work for #CHR2024, we've moved on from the influence of language and translation on stylometric attribution accuracy to the influence of corpus composition. I'll be presenting both parts, but for the corpus composition issue, we are still at the stage of preliminary results.

    With a special thanks to @artjomshl.bsky.social!

    Slides: dhtrier.quarto.pub/icla/

    @rebsim #ICLA2025

  8. My own talk, which is coming up / just happened, is on #Multilingual #Stylometry.

    Based on our work for #CHR2024, we've moved on from the influence of language and translation on stylometric attribution accuracy to the influence of corpus composition. I'll be presenting both parts, but for the corpus composition issue, we are still at the stage of preliminary results.

    With a special thanks to @artjomshl.bsky.social!

    Slides: dhtrier.quarto.pub/icla/

    @rebsim #ICLA2025

  9. Now we're kicking off our "Digital Comparative Literature" track at #ICLA2025 with the first session. Three talks on social reading / Goodreads, on #multilingual #stylometry, and on visualisation of visual data.

    See the session programme here: conftool.pro/icla2025/index.ph

    @rebsim

  10. Now we're kicking off our "Digital Comparative Literature" track at #ICLA2025 with the first session. Three talks on social reading / Goodreads, on #multilingual #stylometry, and on visualisation of visual data.

    See the session programme here: conftool.pro/icla2025/index.ph

    @rebsim

  11. Great to be at the "Comparative Literature Goes Digital" session at #DH2025!

    Session info here: conftool.pro/dh2025/index.php?

    Full programme here: dls.hypotheses.org/1952

    Including a talk by Evgeniia Filveva, with Julia Havrylash, myself, Artjoms Šeļa on "#Multilingual #Stylometry: The influence of corpus composition and language on the performance of authorship attribution using corpora from the European Literary Text Collection (#ELTeC)".

    #ICLA #ADHO #SIG_DLS #CLS @tcdh

  12. Great to be at the "Comparative Literature Goes Digital" session at #DH2025!

    Session info here: conftool.pro/dh2025/index.php?

    Full programme here: dls.hypotheses.org/1952

    Including a talk by Evgeniia Filveva, with Julia Havrylash, myself, Artjoms Šeļa on "#Multilingual #Stylometry: The influence of corpus composition and language on the performance of authorship attribution using corpora from the European Literary Text Collection (#ELTeC)".

    #ICLA #ADHO #SIG_DLS #CLS @tcdh

  13. Reminder for those who may not realize this, but #Stylometry is kind of an insane field of study, and you can be uniquely identified based on your writing style alone.

    This has, in the past, been applied to open source developers and programming code too, and it was found that using stylometry techniques you can identify the author of a
    compiled binary based on their open source code style ~78% of the time

    https://arxiv.org/pdf/1512.08546v1

    There are some techniques to avoid this luckily, which involve fairly basic changes to your writing style and structure that can very effectively anonymize things again:

    https://en.wikipedia.org/wiki/Adversarial_stylometry

  14. Reminder for those who may not realize this, but #Stylometry is kind of an insane field of study, and you can be uniquely identified based on your writing style alone.

    This has, in the past, been applied to open source developers and programming code too, and it was found that using stylometry techniques you can identify the author of a
    compiled binary based on their open source code style ~78% of the time

    https://arxiv.org/pdf/1512.08546v1

    There are some techniques to avoid this luckily, which involve fairly basic changes to your writing style and structure that can very effectively anonymize things again:

    https://en.wikipedia.org/wiki/Adversarial_stylometry

  15. In unserem #StabiLab gibt es Digital Humanities zum Ausprobieren! Am Dienstag, den 21. Januar, lernt ihr bei uns, wie ihr mit dem Tool #Stylo Literatur erforschen könnt 👉 sbb.berlin/59m32

    #StabiBerlin #Stylometry #Stylometrie #Digitalisierung #Forschung #Workshop

  16. In unserem #StabiLab gibt es Digital Humanities zum Ausprobieren! Am Dienstag, den 21. Januar, lernt ihr bei uns, wie ihr mit dem Tool #Stylo Literatur erforschen könnt 👉 sbb.berlin/59m32

    #StabiBerlin #Stylometry #Stylometrie #Digitalisierung #Forschung #Workshop

  17. Later today at #CHR2024, we are going to present our work on #Multilingual #Stylometry!

    We isolated the influence of #language on #authorship #attribution #accuracy by translating multiple #corpora into each others' languages while keeping #corpus composition stable.

    Interactive showcase: showcases.clsinfra.io/stylomet

    Full paper: ceur-ws.org/Vol-3834/paper9.pd

    This work was developed within the @CLSinfra project in #Trier, #Krakow and #Prague with Artjoms Šeļa, Evgeniia Fileva and Julia Dudar.

  18. Later today at #CHR2024, we are going to present our work on #Multilingual #Stylometry!

    We isolated the influence of #language on #authorship #attribution #accuracy by translating multiple #corpora into each others' languages while keeping #corpus composition stable.

    Interactive showcase: showcases.clsinfra.io/stylomet

    Full paper: ceur-ws.org/Vol-3834/paper9.pd

    This work was developed within the @CLSinfra project in #Trier, #Krakow and #Prague with Artjoms Šeļa, Evgeniia Fileva and Julia Dudar.

  19. Agapitos and van Cranenburgh use computational #stylometry to show that while 'Octavia' and 'Hercules Oetaeus' were largely written by #Seneca, a closer analysis of the text segments reveals signs of mixed #authorship. doi.org/10.48694/jcls.3919 #CLS #CCLS24 #Classics #AuthorshipVerification

  20. Agapitos and van Cranenburgh use computational #stylometry to show that while 'Octavia' and 'Hercules Oetaeus' were largely written by #Seneca, a closer analysis of the text segments reveals signs of mixed #authorship. doi.org/10.48694/jcls.3919 #CLS #CCLS24 #Classics #AuthorshipVerification

  21. Look what landed on my doorstep 😍 The book is also available #OpenAccess online at #heiUP: heiup.uni-heidelberg.de/catalo and I would like to thank the very patient editors who had to deal with switching the publisher and coming up with ways to improve the quality of my illustrations in my article about #stylometry in #French and #Spanish for #Picasso 's writings: @christof @josecalvo @u_henny and Robert Hesselbach, Daniel Schlör

  22. Look what landed on my doorstep 😍 The book is also available #OpenAccess online at #heiUP: heiup.uni-heidelberg.de/catalo and I would like to thank the very patient editors who had to deal with switching the publisher and coming up with ways to improve the quality of my illustrations in my article about #stylometry in #French and #Spanish for #Picasso 's writings: @christof @josecalvo @u_henny and Robert Hesselbach, Daniel Schlör

  23. New #paper out: « Code #stylometry vs formatting and minification » peerj.com/articles/cs-2142/ , where we show how much current code stylometry techniques (i.e., how to automatically detect the author of a source code snippet) are resistent to automatic code formatting and minification. (Spoiler: quite a bit, authors can still be identified after those source-to-source transformations.) Available #openaccess on #PeerJ CS.

  24. New #paper out: « Code #stylometry vs formatting and minification » peerj.com/articles/cs-2142/ , where we show how much current code stylometry techniques (i.e., how to automatically detect the author of a source code snippet) are resistent to automatic code formatting and minification. (Spoiler: quite a bit, authors can still be identified after those source-to-source transformations.) Available #openaccess on #PeerJ CS.

  25. Interesting! Dominika Weronska on "A Stylometric Glance at Basque Novels" at #DH2024. #stylometry

    The author did stylometric analyses on 57 Basque novels, a first!

  26. Interesting! Dominika Weronska on "A Stylometric Glance at Basque Novels" at #DH2024. #stylometry

    The author did stylometric analyses on 57 Basque novels, a first!

  27. Now up at #DH2024, Maciej Eder, developer of #stylo and co-organizer of #DH2016 in #Krakow, on various distance measures for #Stylometry: "Manhattan, Euclidean and their Siblings. Exploring Exotic Measures of Text Similarities...".

    Key idea: Manhattan distance is L1-norm based, Euclidean is L2. But we can vary this parameter for a wide range of values, from 0.1 to 10. Then evaluate accuracy for authorship attribution.

    Result: For longer vectors, it pays off to use a value of less than 1!

  28. Now up at #DH2024, Maciej Eder, developer of #stylo and co-organizer of #DH2016 in #Krakow, on various distance measures for #Stylometry: "Manhattan, Euclidean and their Siblings. Exploring Exotic Measures of Text Similarities...".

    Key idea: Manhattan distance is L1-norm based, Euclidean is L2. But we can vary this parameter for a wide range of values, from 0.1 to 10. Then evaluate accuracy for authorship attribution.

    Result: For longer vectors, it pays off to use a value of less than 1!

  29. Kurz mal getestet, stylo() kann die verschiedenen Versformen bei Goethe ziemlich sicher auseinanderhalten: Dramen in Alexandrinern, Knitteln, Blankversen, gemischten Versen sowie die beiden hexametrischen Epen.

    (Volltexte via #DraCor bzw. @gutenberg_org.)

    #DigitalHumanities #Stylometry

  30. Kurz mal getestet, stylo() kann die verschiedenen Versformen bei Goethe ziemlich sicher auseinanderhalten: Dramen in Alexandrinern, Knitteln, Blankversen, gemischten Versen sowie die beiden hexametrischen Epen.

    (Volltexte via #DraCor bzw. @gutenberg_org.)

    #DigitalHumanities #Stylometry

  31. @dvergano … until you start using techniques to defend against #stylometry

    whonix.org/wiki/Stylometry

    (One of the many reasons I love and support the #whonix project)

  32. @dvergano … until you start using techniques to defend against #stylometry

    whonix.org/wiki/Stylometry

    (One of the many reasons I love and support the #whonix project)

  33. @jcls Another paper we would like to highlight, again for the lovers of #novels

    Dorothy Henriette Modrall Sperling, Mike Kestemont & Vincent Neyt (2023), “The Authorship of Stephen King’s Books Written Under the Pseudonym “Richard #Bachman”: A Stylometric Analysis”, Journal of Computational Literary Studies 2(1), 1–35. doi: doi.org/10.48694/jcls.3594

    Keywords: #Stephen_King, #stylometry, #pop_culture, #authorship verification, contemporary English-language #fiction

  34. @jcls Another paper we would like to highlight, again for the lovers of #novels

    Dorothy Henriette Modrall Sperling, Mike Kestemont & Vincent Neyt (2023), “The Authorship of Stephen King’s Books Written Under the Pseudonym “Richard #Bachman”: A Stylometric Analysis”, Journal of Computational Literary Studies 2(1), 1–35. doi: doi.org/10.48694/jcls.3594

    Keywords: #Stephen_King, #stylometry, #pop_culture, #authorship verification, contemporary English-language #fiction

  35. This next paper is about #stylometry in a #translation setting involving novels in #Swedish and #Danish:

    Martje Wijers (2023), “Why the Daisy sisters are different. A stylometric study on the oeuvre of Swedish author Henning #Mankell and the Dutch translations of his work”, Journal of Computational Literary Studies 2 (1), 1–27. doi: doi.org/10.48694/jcls.3585

    Keywords: #stylometry, #cluster analysis, #PCA, #delta, #zeta, #translation

  36. This next paper is about #stylometry in a #translation setting involving novels in #Swedish and #Danish:

    Martje Wijers (2023), “Why the Daisy sisters are different. A stylometric study on the oeuvre of Swedish author Henning #Mankell and the Dutch translations of his work”, Journal of Computational Literary Studies 2 (1), 1–27. doi: doi.org/10.48694/jcls.3585

    Keywords: #stylometry, #cluster analysis, #PCA, #delta, #zeta, #translation

  37. Very happy to participate in today's workshop on "Potentials and Limits of #Stylometry for Early Modern Text in #Romance Languages". It's co-organized by the "Pamphlets and Patrons" #PAPA project in Early Modern French History and the Trier Center for Digital Humanities @tcdh today.

    The programme is here: tcdh.uni-trier.de/en/event/hyb

    #CLS #Romanistik #Trier

  38. The team of project REWIND invites you to register for the workshop “In search of the gender signal: introduction to stylometric approaches”, which we will host on 18 September.

    Conducted by Helena Bermúdez Sabel, it will explore stylometric methods to tease out gender makers in literary corpora in Romance languages.

    Hybrid. FREE REGISTRATION

    ℹ️ ihc.fcsh.unl.pt/en/events/sear

    @histodons
    @litstudies
    #histodons #litstudies #DigitalHumanities #Stylometry #GenderStudies #Literature