64: A Garbled PDF episode artwork

EPISODE · Jul 3, 2025 · 23 MIN

64: A Garbled PDF

from Base by Base · host Gustavo Barra

Xu H et al., Cell Genomics - This episode examines a heavily corrupted PDF provided as the source. The text is dominated by recurring, unreadable tokens (e.g., Wt�mo�, m{yltzk�t{z, k�ryoz�k�t{z) and fragmented sections, preventing clear extraction of aims or results. We walk listeners through what can and cannot be recovered from the file. Key terms: Wt�mo�, m{yltzk�t{z, k�ryoz�k�t{z, oqqom�t�owÞ, J~�rGkzv. Study Highlights:The supplied PDF is extensively corrupted and repeatedly contains tokens such as "Wt�mo�", "m{yltzk�t{z", and "k�ryoz�k�t{z" that recur throughout. Sections also reference forms like "oqqom�t�owÞ" and labels such as "J~�rGkzv", suggesting structured headings or entities but unreadable encoding. Because of pervasive formatting and encoding errors the study's aims, methods and results cannot be reliably extracted from the text. Conclusion:The PDF text is too corrupted to recover definitive conclusions; a clean source is required for meaningful interpretation. Music:Enjoy the music based on this article at the end of the episode. Article title:Pisces: A multi-modal data augmentation approach for drug combination synergy prediction First author:Xu H Journal:Cell Genomics DOI:10.1016/j.xgen.2025.100892 Reference:Xu H., Lin J., Woicik A., Liu Z., Ma J., Zhang S., et al.. Pisces: A multi-modal data augmentation approach for drug combination synergy prediction. Cell Genomics, 5, 100892. (2025). https://doi.org/10.1016/j.xgen.2025.100892 License:This episode is based on an open-access article published under the Creative Commons Attribution 4.0 International License (CC BY 4.0) – https://creativecommons.org/licenses/by/4.0/ Support:Base by Base – Stripe donations: https://donate.stripe.com/7sY4gz71B2sN3RWac5gEg00 Official website https://basebybase.com On PaperCast Base by Base you'll discover the latest in genomics, functional genomics, structural genomics, and proteomics. Episode link: https://basebybase.com/episodes/base-by-base-64-garbled-pdf QC:This episode was checked against the original article PDF and publication metadata for the episode release published on 2025-07-03. QC Scope:- article metadata and core scientific claims from the narration- excludes analogies, intro/outro, and music- transcript coverage: Substantively audited sections describing Pisces architecture, data augmentation, the 64-view augmenter, the noisy-label aggregator, and the key experimental results (cell lines, unseen drug pairs, 3-drug synergy, in vivo) plus limitations and clinical implications.- transcript topics: Problem of drug synergy and data scarcity; Multimodal data augmentation concept (Pisces); Eight modalities per drug and universal embedding; The augmenter: 8 x 8 views = 64 augmented views; Noisy label aggregator selecting top 8 predictions; Evaluation on GDSC data and unseen drug pair/cell line splits QC Summary:- factual score: 10/10- metadata score: 10/10- supported core claims: 7- claims flagged for review: 0- metadata checks passed: 4- metadata issues found: 0 Metadata Audited:- article_doi- article_title- article_journal- license Factual Items Audited:- Pisces uses 8 modalities per drug and forms 64 augmented views for each drug pair- Projector translates cross-modality representations into a shared embedding space- Aggregator employs noisy label learning and retains the top 8 predictions- Unseen 2-drug combinations: F1 improves by ~24% over the next-best approach- Unseen cell lines: F1 improvement > ~10%- Triplet (3-drug) synergy evaluation: AUROC = 0.8525 QC result: Pass.

Xu H et al., Cell Genomics - This episode examines a heavily corrupted PDF provided as the source. The text is dominated by recurring, unreadable tokens (e.g., Wt�mo�, m{yltzk�t{z, k�ryoz�k�t{z) and fragmented sections, preventing clear extraction of aims or results. We walk listeners through what can and cannot be recovered from the file. Key terms: Wt�mo�, m{yltzk�t{z, k�ryoz�k�t{z, oqqom�t�owÞ, J~�rGkzv. Study Highlights:The supplied PDF is extensively corrupted and repeatedly contains tokens such as "Wt�mo�", "m{yltzk�t{z", and "k�ryoz�k�t{z" that recur throughout. Sections also reference forms like "oqqom�t�owÞ" and labels such as "J~�rGkzv", suggesting structured headings or entities but unreadable encoding. Because of pervasive formatting and encoding errors the study's aims, methods and results cannot be reliably extracted from the text. Conclusion:The PDF text is too corrupted to recover definitive conclusions; a clean source is required for meaningful interpretation. Music:Enjoy the music based on this article at the end of the episode. Article title:Pisces: A multi-modal data augmentation approach for drug combination synergy prediction First author:Xu H Journal:Cell Genomics DOI:10.1016/j.xgen.2025.100892 Reference:Xu H., Lin J., Woicik A., Liu Z., Ma J., Zhang S., et al.. Pisces: A multi-modal data augmentation approach for drug combination synergy prediction. Cell Genomics, 5, 100892. (2025). https://doi.org/10.1016/j.xgen.2025.100892 License:This episode is based on an open-access article published under the Creative Commons Attribution 4.0 International License (CC BY 4.0) – https://creativecommons.org/licenses/by/4.0/ Support:Base by Base – Stripe donations: https://donate.stripe.com/7sY4gz71B2sN3RWac5gEg00 Official website https://basebybase.com On PaperCast Base by Base you'll discover the latest in genomics, functional genomics, structural genomics, and proteomics. Episode link: https://basebybase.com/episodes/base-by-base-64-garbled-pdf QC:This episode was checked against the original article PDF and publication metadata for the episode release published on 2025-07-03. QC Scope:- article metadata and core scientific claims from the narration- excludes analogies, intro/outro, and music- transcript coverage: Substantively audited sections describing Pisces architecture, data augmentation, the 64-view augmenter, the noisy-label aggregator, and the key experimental results (cell lines, unseen drug pairs, 3-drug synergy, in vivo) plus limitations and clinical implications.- transcript topics: Problem of drug synergy and data scarcity; Multimodal data augmentation concept (Pisces); Eight modalities per drug and universal embedding; The augmenter: 8 x 8 views = 64 augmented views; Noisy label aggregator selecting top 8 predictions; Evaluation on GDSC data and unseen drug pair/cell line splits QC Summary:- factual score: 10/10- metadata score: 10/10- supported core claims: 7- claims flagged for review: 0- metadata checks passed: 4- metadata issues found: 0 Metadata Audited:- article_doi- article_title- article_journal- license Factual Items Audited:- Pisces uses 8 modalities per drug and forms 64 augmented views for each drug pair- Projector translates cross-modality representations into a shared embedding space- Aggregator employs noisy label learning and retains the top 8 predictions- Unseen 2-drug combinations: F1 improves by ~24% over the next-best approach- Unseen cell lines: F1 improvement > ~10%- Triplet (3-drug) synergy evaluation: AUROC = 0.8525 QC result: Pass.

NOW PLAYING

64: A Garbled PDF

0:00 23:10

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

MG Show MG Show The MG Show, hosted by Jeffrey Pedersen and Shannon Townsend, is a leading alternative media platform dedicated to uncovering the truth behind today’s most pressing political issues. Launched in 2019, the show has grown exponentially, offering unfiltered insights, comprehensive research, and real-time analysis. With a commitment to independent journalism and factual integrity, the MG Show empowers its audience with knowledge and encourages active participation in the political discourse. That Hoarder: Overcome Compulsive Hoarding That Hoarder Hoarding disorder is stigmatised and people who hoard feel vast amounts of shame. This podcast began life as an audio diary, an anonymous outlet for somebody with this weird condition. That Hoarder speaks about her experiences living with compulsive hoarding, she interviews therapists, academics, researchers, children of hoarders, professional organisers and influencers, and she shares insight and tips for others with the problem. Listened to by people who hoard as well as those who love them and those who work with them, Overcome Compulsive Hoarding with That Hoarder aims to shatter the stigma, share the truth and speak openly and honestly to improve lives. Flottengeflüster ALD Automotive Österreich | LeasePlan Beim Flottengeflüster powered by ALD Automotive | LeasePlan präsentieren Jörg Janik und Peter Gutenbrunner alle zwei Wochen spannende Informationen rund um das Thema nachhaltige Mobilität. Beide beschäftigen sich schon lange mit der Thematik und bringen umfangreiches Fachwissen mit. Sollten sie aber doch einmal nicht weiter wissen, werden unsere Expert*innen hinzugezogen, die ihnen gerne mit Rat und Tat zur Seite stehen. The Small Business Startup School – Business Notes | Financial Literacy | Retail Psychology – For Professionals & Entrepreneurs The Small Business Startup School Inc. Starting or buying a small business? While personal circumstances may vary, business patterns remain timeless. On The Small Business Startup School, we explore strategies, insights, and practical solutions to help entrepreneurs confidently navigate their journey.Hosted by Ola Williams—a retail entrepreneur, fintech founder, and financial coach with over two decades of experience—this podcast marries financial awareness and retail psychology with optimism to deliver actionable takeaways.Join us to learn, grow, and connect as we uncover the keys to business success.Let’s continue to learn together and be encouraged to keep on connecting!

Frequently Asked Questions

How long is this episode of Base by Base?

This episode is 23 minutes long.

When was this Base by Base episode published?

This episode was published on July 3, 2025.

What is this episode about?

Xu H et al., Cell Genomics - This episode examines a heavily corrupted PDF provided as the source. The text is dominated by recurring, unreadable tokens (e.g., Wt�mo�, m{yltzk�t{z, k�ryoz�k�t{z) and fragmented sections, preventing clear extraction...

Can I download this Base by Base episode?

Yes, you can download this episode by clicking the download button on the episode player, or subscribe to the podcast in your preferred podcast app for automatic downloads.
URL copied to clipboard!