Rogue Scholar Posts

language
Published in GigaBlog

TDR, the Special Programme for Research and Training in Tropical Diseases hosted at the World Health Organization (WHO), GBIF and GigaScience Press have announced a third call for authors to submit Data Release papers on vectors of human disease for inclusion in a thematic series published in GigaByte Journal.

Published in bjoern.brembs.blog

Universities worldwide currently face a pivotal choice: should they contribute to building a global infrastructure for exchange, science, and discourse, free from the control of oligarchs, to promote democracy, human rights, and digital participation? Or should they continue advertising on private networks, hoping for clicks and marginally increased student enrollment?

Published in Paired Ends
Author Stephen Turner

This week’s recap highlights the Evo model for sequence modeling and design, biomedical discovery with AI agents, improving bioinformatics software quality through teamwork, a new tool from Brent Pedersen and Aaron Quinlan (vcfexpress) for filtering and formatting VCFs with Lua expressions, a new paper about the NHGRI-EBI GWAS Catalog, and a review paper on designing and engineering synthetic genomes.

Published in Paired Ends
Author Stephen Turner

A few days ago I wrote about translating R package help documentation using a local LLM (e.g. llama3.x)… …when Mick Watson commented: I was already thinking of wiring up something like this using local AI models — something to summarize podcasts, conference recordings, etc. The relatively new (as of this writing) Gemini 2.0 Flash model will do this for you for YouTube videos. But what if you wanted to do this offline using a local LLM?

Published in Paired Ends
Author Stephen Turner

Last week I posted about a web app that turns a GitHub repo into a single text file for LLM-friendly input. This is great for capturing LLM-friendly text from a GitHub repo, but what about any other arbitrary website or PDF? I was catching up on Simon Willison’s newsletter reading about an app he made with Claude artifacts that uses the Jina Reader API to generate Markdown from a website. You don’t need to use the API to do this.

Published in Paired Ends
Author Stephen Turner

Using LLMs in R Most of the developer tooling for AI/LLM training and evaluation is Python-centric, but just over the past few months we’ve seen a surge of new tooling for AI/LLM applications for the R ecosystem. ollamar and rollama provide wrappers around the Ollama API allowing you to run LLMs locally on your machine.

Published in Paired Ends
Author Stephen Turner

This week’s recap highlights a new way to turn Nextflow pipelines into web apps, DRAGEN for fast and accurate variant calling, machine-guided design of cell-type-targeting cis-regulatory elements, a Nextflow pipeline for identifying and classifying protein kinases, a new language model for single cell perturbations that integrates knowledge from literature, GeneCards, etc., and a new method for scalable protein design in a relaxed sequence