home.social

#ismbeccb2023 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ismbeccb2023, aggregated by home.social.

fetched live
  1. We had an excellent #CompMS session at the #ISMBECCB2023 conference last week.

    Many thanks to keynote speakers @[email protected], @[email protected], and @[email protected]; all selected speakers; and poster presenters for showcasing the latest computational advances in mass spectrometry, with applications across #proteomics, #metabolomics, #lipidomics, and more.

  2. We had an excellent #CompMS session at the #ISMBECCB2023 conference last week.

    Many thanks to keynote speakers @[email protected], @[email protected], and @[email protected]; all selected speakers; and poster presenters for showcasing the latest computational advances in mass spectrometry, with applications across #proteomics, #metabolomics, #lipidomics, and more.

  3. We had an excellent #CompMS session at the #ISMBECCB2023 conference last week.

    Many thanks to keynote speakers @[email protected], @[email protected], and @[email protected]; all selected speakers; and poster presenters for showcasing the latest computational advances in mass spectrometry, with applications across #proteomics, #metabolomics, #lipidomics, and more.

  4. We had an excellent #CompMS session at the #ISMBECCB2023 conference last week.

    Many thanks to keynote speakers @[email protected], @[email protected], and @[email protected]; all selected speakers; and poster presenters for showcasing the latest computational advances in mass spectrometry, with applications across #proteomics, #metabolomics, #lipidomics, and more.

  5. We had an excellent #CompMS session at the #ISMBECCB2023 conference last week.

    Many thanks to keynote speakers @[email protected], @[email protected], and @[email protected]; all selected speakers; and poster presenters for showcasing the latest computational advances in mass spectrometry, with applications across #proteomics, #metabolomics, #lipidomics, and more.

  6. RT @: Keep calm, Pfam is still running!But now it's hosted on the InterPro website! At #ISMBECCB2023, we had the opportunity to learn more about @PfamDB and its integration with @InterProDB website. We even won these really cool t-shirts,Thanks!

  7. RT @: Keep calm, Pfam is still running!But now it's hosted on the InterPro website! At #ISMBECCB2023, we had the opportunity to learn more about @PfamDB and its integration with @InterProDB website. We even won these really cool t-shirts,Thanks!

  8. RT @: Keep calm, Pfam is still running!But now it's hosted on the InterPro website! At #ISMBECCB2023, we had the opportunity to learn more about @PfamDB and its integration with @InterProDB website. We even won these really cool t-shirts,Thanks!

  9. RT @: Keep calm, Pfam is still running!But now it's hosted on the InterPro website! At #ISMBECCB2023, we had the opportunity to learn more about @PfamDB and its integration with @InterProDB website. We even won these really cool t-shirts,Thanks!

  10. RT @: Keep calm, Pfam is still running!But now it's hosted on the InterPro website! At #ISMBECCB2023, we had the opportunity to learn more about @PfamDB and its integration with @InterProDB website. We even won these really cool t-shirts,Thanks!

  11. Mark Gerstein at #ISMBECCB2023: Deep learning is exciting, but let's not forget about the physical and biological models underlying the science we're interested in. Let's make biomedical data science more like weather forecasting.

  12. Mark Gerstein at #ISMBECCB2023: Deep learning is exciting, but let's not forget about the physical and biological models underlying the science we're interested in. Let's make biomedical data science more like weather forecasting.

  13. Mark Gerstein at #ISMBECCB2023: Deep learning is exciting, but let's not forget about the physical and biological models underlying the science we're interested in. Let's make biomedical data science more like weather forecasting.

  14. Mark Gerstein at #ISMBECCB2023: Deep learning is exciting, but let's not forget about the physical and biological models underlying the science we're interested in. Let's make biomedical data science more like weather forecasting.

  15. Mark Gerstein at #ISMBECCB2023: Deep learning is exciting, but let's not forget about the physical and biological models underlying the science we're interested in. Let's make biomedical data science more like weather forecasting.

  16. Névéol: What can we do?
    Understand the stakes better.
    Facilitate levers like data sharing, shared tasks, and policy.
    Write more documentation, for protocols, etc.; elicit audits.

    See Cohen-Boulakia et al 2017 Future Gen Comput Syst

    #ismbeccb2023
    #textmining

  17. Névéol: What can we do?
    Understand the stakes better.
    Facilitate levers like data sharing, shared tasks, and policy.
    Write more documentation, for protocols, etc.; elicit audits.

    See Cohen-Boulakia et al 2017 Future Gen Comput Syst

    #ismbeccb2023
    #textmining

  18. Névéol: What can we do?
    Understand the stakes better.
    Facilitate levers like data sharing, shared tasks, and policy.
    Write more documentation, for protocols, etc.; elicit audits.

    See Cohen-Boulakia et al 2017 Future Gen Comput Syst

    #ismbeccb2023
    #textmining

  19. Névéol: What can we do?
    Understand the stakes better.
    Facilitate levers like data sharing, shared tasks, and policy.
    Write more documentation, for protocols, etc.; elicit audits.

    See Cohen-Boulakia et al 2017 Future Gen Comput Syst

    #ismbeccb2023
    #textmining

  20. Névéol: What can we do?
    Understand the stakes better.
    Facilitate levers like data sharing, shared tasks, and policy.
    Write more documentation, for protocols, etc.; elicit audits.

    See Cohen-Boulakia et al 2017 Future Gen Comput Syst

    #ismbeccb2023
    #textmining

  21. Aurélie Névéol:
    How can we make clinical NLP more reproducible? Can NLP also help with reproducibility? Even word or sentence tokenization can be inconsistent. Most NLP folks have, at least once, failed to repeat someone else's experiment, or even their own. Sometimes it's due to differences in preprocessing, software versions, training vs test splits, or other boring things. Availability issues, page limits, and the bias toward novelty don't help either.

    #ismbeccb2023
    #textmining

  22. Aurélie Névéol:
    How can we make clinical NLP more reproducible? Can NLP also help with reproducibility? Even word or sentence tokenization can be inconsistent. Most NLP folks have, at least once, failed to repeat someone else's experiment, or even their own. Sometimes it's due to differences in preprocessing, software versions, training vs test splits, or other boring things. Availability issues, page limits, and the bias toward novelty don't help either.

    #ismbeccb2023
    #textmining

  23. Aurélie Névéol:
    How can we make clinical NLP more reproducible? Can NLP also help with reproducibility? Even word or sentence tokenization can be inconsistent. Most NLP folks have, at least once, failed to repeat someone else's experiment, or even their own. Sometimes it's due to differences in preprocessing, software versions, training vs test splits, or other boring things. Availability issues, page limits, and the bias toward novelty don't help either.

    #ismbeccb2023
    #textmining

  24. Aurélie Névéol:
    How can we make clinical NLP more reproducible? Can NLP also help with reproducibility? Even word or sentence tokenization can be inconsistent. Most NLP folks have, at least once, failed to repeat someone else's experiment, or even their own. Sometimes it's due to differences in preprocessing, software versions, training vs test splits, or other boring things. Availability issues, page limits, and the bias toward novelty don't help either.

    #ismbeccb2023
    #textmining

  25. Aurélie Névéol:
    How can we make clinical NLP more reproducible? Can NLP also help with reproducibility? Even word or sentence tokenization can be inconsistent. Most NLP folks have, at least once, failed to repeat someone else's experiment, or even their own. Sometimes it's due to differences in preprocessing, software versions, training vs test splits, or other boring things. Availability issues, page limits, and the bias toward novelty don't help either.

    #ismbeccb2023
    #textmining

  26. One perk of attending #ISMBECCB2023 virtually: watching the recording of a keynote I missed instead of the talk I had planned to watch but turned out not to be interested in.

    (I guess you could also plug in your headphones and do the same if you're there in person, but that's noticeably ruder.)

  27. One perk of attending #ISMBECCB2023 virtually: watching the recording of a keynote I missed instead of the talk I had planned to watch but turned out not to be interested in.

    (I guess you could also plug in your headphones and do the same if you're there in person, but that's noticeably ruder.)

  28. One perk of attending #ISMBECCB2023 virtually: watching the recording of a keynote I missed instead of the talk I had planned to watch but turned out not to be interested in.

    (I guess you could also plug in your headphones and do the same if you're there in person, but that's noticeably ruder.)

  29. One perk of attending #ISMBECCB2023 virtually: watching the recording of a keynote I missed instead of the talk I had planned to watch but turned out not to be interested in.

    (I guess you could also plug in your headphones and do the same if you're there in person, but that's noticeably ruder.)

  30. One perk of attending #ISMBECCB2023 virtually: watching the recording of a keynote I missed instead of the talk I had planned to watch but turned out not to be interested in.

    (I guess you could also plug in your headphones and do the same if you're there in person, but that's noticeably ruder.)

  31. Sylwia Szymanska: Word embeddings capture functions of low complexity regions: scientific literature analysis using a transformer-based language model

    Low-complexity regions in proteins are biologically important. But there isn't a database or even a list of these relationships. So let's extract them with a language model.
    #ismbeccb2023
    #textmining

  32. Sylwia Szymanska: Word embeddings capture functions of low complexity regions: scientific literature analysis using a transformer-based language model

    Low-complexity regions in proteins are biologically important. But there isn't a database or even a list of these relationships. So let's extract them with a language model.
    #ismbeccb2023
    #textmining

  33. Sylwia Szymanska: Word embeddings capture functions of low complexity regions: scientific literature analysis using a transformer-based language model

    Low-complexity regions in proteins are biologically important. But there isn't a database or even a list of these relationships. So let's extract them with a language model.
    #ismbeccb2023
    #textmining

  34. Sylwia Szymanska: Word embeddings capture functions of low complexity regions: scientific literature analysis using a transformer-based language model

    Low-complexity regions in proteins are biologically important. But there isn't a database or even a list of these relationships. So let's extract them with a language model.
    #ismbeccb2023
    #textmining

  35. Sylwia Szymanska: Word embeddings capture functions of low complexity regions: scientific literature analysis using a transformer-based language model

    Low-complexity regions in proteins are biologically important. But there isn't a database or even a list of these relationships. So let's extract them with a language model.
    #ismbeccb2023
    #textmining

  36. Brett Beaulieu-Jones: Can we use large language models with clinical notes to estimate likelihood of seizure recurrence? Yes - and even with good results - but models are difficult to interpret. So can we build a model that includes things we really care about, then add an instructable layer? Yes! Use note metadata as weak supervision -> instructions for the model. A tuned T5-Flan model does really well.

    #ismbeccb2023
    #textmining

  37. Brett Beaulieu-Jones: Can we use large language models with clinical notes to estimate likelihood of seizure recurrence? Yes - and even with good results - but models are difficult to interpret. So can we build a model that includes things we really care about, then add an instructable layer? Yes! Use note metadata as weak supervision -> instructions for the model. A tuned T5-Flan model does really well.

    #ismbeccb2023
    #textmining

  38. Brett Beaulieu-Jones: Can we use large language models with clinical notes to estimate likelihood of seizure recurrence? Yes - and even with good results - but models are difficult to interpret. So can we build a model that includes things we really care about, then add an instructable layer? Yes! Use note metadata as weak supervision -> instructions for the model. A tuned T5-Flan model does really well.

    #ismbeccb2023
    #textmining

  39. Brett Beaulieu-Jones: Can we use large language models with clinical notes to estimate likelihood of seizure recurrence? Yes - and even with good results - but models are difficult to interpret. So can we build a model that includes things we really care about, then add an instructable layer? Yes! Use note metadata as weak supervision -> instructions for the model. A tuned T5-Flan model does really well.

    #ismbeccb2023
    #textmining

  40. Brett Beaulieu-Jones: Can we use large language models with clinical notes to estimate likelihood of seizure recurrence? Yes - and even with good results - but models are difficult to interpret. So can we build a model that includes things we really care about, then add an instructable layer? Yes! Use note metadata as weak supervision -> instructions for the model. A tuned T5-Flan model does really well.

    #ismbeccb2023
    #textmining

  41. Robert Leaman: BioNER requires multiple entity types -> relation extraction. But human-annotated NER data are scarce - < 0.01% of PubMed articles. Can use pre trained language model to do multitask NER...but instead we could modify the data. Include annotations for negative mentions by type + tokens for sentence start/end.
    AIONER converts training data to this form and aggregates data sets. Moderate improvement on most BioNER types.

    Repo here: github.com/ncbi/AIONER

    #ismbeccb2023
    #textmining

  42. Robert Leaman: BioNER requires multiple entity types -> relation extraction. But human-annotated NER data are scarce - < 0.01% of PubMed articles. Can use pre trained language model to do multitask NER...but instead we could modify the data. Include annotations for negative mentions by type + tokens for sentence start/end.
    AIONER converts training data to this form and aggregates data sets. Moderate improvement on most BioNER types.

    Repo here: github.com/ncbi/AIONER

    #ismbeccb2023
    #textmining

  43. Robert Leaman: BioNER requires multiple entity types -> relation extraction. But human-annotated NER data are scarce - < 0.01% of PubMed articles. Can use pre trained language model to do multitask NER...but instead we could modify the data. Include annotations for negative mentions by type + tokens for sentence start/end.
    AIONER converts training data to this form and aggregates data sets. Moderate improvement on most BioNER types.

    Repo here: github.com/ncbi/AIONER

    #ismbeccb2023
    #textmining

  44. Robert Leaman: BioNER requires multiple entity types -> relation extraction. But human-annotated NER data are scarce - < 0.01% of PubMed articles. Can use pre trained language model to do multitask NER...but instead we could modify the data. Include annotations for negative mentions by type + tokens for sentence start/end.
    AIONER converts training data to this form and aggregates data sets. Moderate improvement on most BioNER types.

    Repo here: github.com/ncbi/AIONER

    #ismbeccb2023
    #textmining

  45. Robert Leaman: BioNER requires multiple entity types -> relation extraction. But human-annotated NER data are scarce - < 0.01% of PubMed articles. Can use pre trained language model to do multitask NER...but instead we could modify the data. Include annotations for negative mentions by type + tokens for sentence start/end.
    AIONER converts training data to this form and aggregates data sets. Moderate improvement on most BioNER types.

    Repo here: github.com/ncbi/AIONER

    #ismbeccb2023
    #textmining

  46. Katerina Nastou: Benchmarking species name NER is some thing the S800 corpus was used for, but otherwise well-performing models were doing poorly on it. Problem? Annotation inconsistencies in S800. It's been manually revised using stricter rules, and just species, strain, and genera names (each with their own tags). 200 more documents too, so now it's S1000.

    How do NER models do on it? F1 up around 89 to 91.

    Get corpus at zenodo.org/record/7064902

    #ismbeccb2023
    #textmining

  47. Katerina Nastou: Benchmarking species name NER is some thing the S800 corpus was used for, but otherwise well-performing models were doing poorly on it. Problem? Annotation inconsistencies in S800. It's been manually revised using stricter rules, and just species, strain, and genera names (each with their own tags). 200 more documents too, so now it's S1000.

    How do NER models do on it? F1 up around 89 to 91.

    Get corpus at zenodo.org/record/7064902

    #ismbeccb2023
    #textmining

  48. Katerina Nastou: Benchmarking species name NER is some thing the S800 corpus was used for, but otherwise well-performing models were doing poorly on it. Problem? Annotation inconsistencies in S800. It's been manually revised using stricter rules, and just species, strain, and genera names (each with their own tags). 200 more documents too, so now it's S1000.

    How do NER models do on it? F1 up around 89 to 91.

    Get corpus at zenodo.org/record/7064902

    #ismbeccb2023
    #textmining

  49. Katerina Nastou: Benchmarking species name NER is some thing the S800 corpus was used for, but otherwise well-performing models were doing poorly on it. Problem? Annotation inconsistencies in S800. It's been manually revised using stricter rules, and just species, strain, and genera names (each with their own tags). 200 more documents too, so now it's S1000.

    How do NER models do on it? F1 up around 89 to 91.

    Get corpus at zenodo.org/record/7064902

    #ismbeccb2023
    #textmining

  50. Katerina Nastou: Benchmarking species name NER is some thing the S800 corpus was used for, but otherwise well-performing models were doing poorly on it. Problem? Annotation inconsistencies in S800. It's been manually revised using stricter rules, and just species, strain, and genera names (each with their own tags). 200 more documents too, so now it's S1000.

    How do NER models do on it? F1 up around 89 to 91.

    Get corpus at zenodo.org/record/7064902

    #ismbeccb2023
    #textmining

  51. Esmaeil Nourani: Health involves lifestyle factors. Can we extract relations connecting those to disease?
    Developing a draft lifestyle ontology. Started with 869 concepts across multiple branches. Needed to get synonyms, too - embeddings helped with that, and also allowed discovery of new candidate terms. Full draft is now 1652 concepts. Ready for NER and RE.

    #ismbeccb2023
    #textmining

  52. Esmaeil Nourani: Health involves lifestyle factors. Can we extract relations connecting those to disease?
    Developing a draft lifestyle ontology. Started with 869 concepts across multiple branches. Needed to get synonyms, too - embeddings helped with that, and also allowed discovery of new candidate terms. Full draft is now 1652 concepts. Ready for NER and RE.

    #ismbeccb2023
    #textmining

  53. Esmaeil Nourani: Health involves lifestyle factors. Can we extract relations connecting those to disease?
    Developing a draft lifestyle ontology. Started with 869 concepts across multiple branches. Needed to get synonyms, too - embeddings helped with that, and also allowed discovery of new candidate terms. Full draft is now 1652 concepts. Ready for NER and RE.

    #ismbeccb2023
    #textmining

  54. Esmaeil Nourani: Health involves lifestyle factors. Can we extract relations connecting those to disease?
    Developing a draft lifestyle ontology. Started with 869 concepts across multiple branches. Needed to get synonyms, too - embeddings helped with that, and also allowed discovery of new candidate terms. Full draft is now 1652 concepts. Ready for NER and RE.

    #ismbeccb2023
    #textmining

  55. Esmaeil Nourani: Health involves lifestyle factors. Can we extract relations connecting those to disease?
    Developing a draft lifestyle ontology. Started with 869 concepts across multiple branches. Needed to get synonyms, too - embeddings helped with that, and also allowed discovery of new candidate terms. Full draft is now 1652 concepts. Ready for NER and RE.

    #ismbeccb2023
    #textmining