No. 024

Rosalind Workbench, Claude for scientists, LLM edits alter meaning

Google released Co-Scientist last month. Since then, Anthropic opened 10,000 subsidized seats for researchers and OpenAI launched a life-sciences workbench with its own domain model. All three frontier labs now sell a scientist-branded product line, making this a product category rather than a series of one-offs. What none of the product pages address is that grammar-only LLM edits shift content and meaning. A team from Princeton and UCSD reported a nearly 70 percent increase in neutral essay responses when they applied AI edits marketed as purely mechanical.

Scientist Products

  • Meet Rosalind Workbench: Empowering every scientist to be their own research team

    OpenAI, August 28 2026

    A dedicated life-sciences environment inside ChatGPT, built on GPT-Rosalind (a domain-specific model for biochemistry, genomics, and chemistry), with molecular viewers, sequence alignment tools, and provenance tracking attached to each result, though the entire system runs within a single closed vendor.

  • Claude Team plan for scientists

    Anthropic, August 27 2026

    Free standard seats and $15/month premium seats for up to 10,000 PIs at nonprofit research institutions, bundled with Claude Science (launched June 30), Code, and connectors to PubMed, bioRxiv, and ChEMBL, with eligibility restricted to academic and nonprofit labs and excluding industry R&D.

  • freephdlabor: open-source multiagent framework for autonomous scientific research

    GitHub (ltjed), August 2026

    An MIT-licensed multiagent framework that orchestrates specialized agents (ideation, writeup, experiments) to produce papers end-to-end, positioning itself as the open-source counterpart to the frontier-lab bundles now forming a category.

Research Practice

  • How do LLMs affect writing? Talks by visitors to the AI STORIES project

    Center for Digital Narrative, University of Bergen, September 4 2026

    Five talks on measured effects of LLMs on research writing, with the key finding from Abdulhai (Princeton) and White (UCSD/Microsoft Research) that grammar-only LLM edits cause a nearly 70 percent increase in neutral essay responses, meaning what looks like a light copyedit is actually a content change.

  • 98 out of 100: humans solved one of the two math problems AI could not, at a Leipzig benchmark

    Czech Technical University in Prague, August 2026

    At a Leipzig advanced mathematics benchmark, AI models solved 98 of 100 problems, and a team from CTU Prague and Universität Leipzig cracked one of the two remaining (a combinatorics case involving breaking triples with permutations), a useful contrast to the 30 percent pass rate on open-ended research tasks from Terminal-Bench Science.

  • AI-Assisted Tools in the CHI 2027 Papers Review Process

    ACM CHI 2027 organizing committee, August 29 2026

    CHI 2027 will deploy AI for submission completeness checks, reviewer matching, and desk-reject screening, but keeps humans responsible for every acceptance decision and deliberately shelved a rubric tool that would have scored paper content, drawing a clear line between automation for logistics and automation for judgment.