(Roughly) Daily

Posts Tagged ‘scientific research

“Science is a cooperative enterprise spanning the generations… a community of minds, reaching back to antiquity and forward to the stars”*…

The sharing of experimental results and the underlying data is critical to the advance of science. Indeed, when I had the chance to do a scenario planning exercise with a collection of the leading research university librarians in the U.S. a couple of decades ago, the biggest threat/fear they surfaced was the concern that the free and open exchange of ideas and data, as manifest formally in scientific publication and informally in the collegial cooperation among scientists, would be occluded by an increasing proprietary embrace of knowledge.

73% of geneticists surveyed in an article in the 23/30 January 2002 issue of the Journal of the American Medical Association agreed that although keeping data private may help the individual researcher, data hoarding is detrimental to the progress of science Still, sadly, that threat has grown since the turn of the millennium.

By way of current (and dramatic) example: as Celina Zhao reports, more than half of AI “unicorns” have never published a paper or preprint…

Today’s biggest artificial intelligence (AI) startups make no shortage of bold promises. Their technologies, some boast, will revolutionize software development, drug discovery, and scientific research.

Yet a new preprint posted on 16 July on bioRxiv suggests many of these firms barely participate in one of science’s most fundamental practices: publicly documenting discoveries in scientific literature so other researchers can evaluate and build on them. More than half of AI unicorns—private companies valued at more than $1 billion—have never played a leading role in publishing a scientific paper or preprint, according to the new analysis. Collectively, they accounted for just one in every 1000 AI papers published in 2025.

“For a field that is supposedly reshaping science and is so advanced in terms of scientific potential, not having any scientific documentation seems like a very weird paradox,” says paper co-author John Ioannidis, a metascientist at Stanford University [see here]. “How can you judge that what they say is real, validated, and reproducible?” The scarcity of publications, others say, also makes it harder to assess AI’s social impacts, including energy use and safety.

But University of Alberta AI ethicist Mohamed Abdalla says the findings reflect the incentives facing commercial AI developers, rather than solely a failure to uphold scientific norms. “It’s not the company’s job to advance science, right?” he says. “The company’s job is to advance money.”

Ioannidis has long studied how unicorns, particularly in biotech, engage with the scientific literature. (In 2015, he was the first to publicly scrutinize the lack of peer-reviewed studies produced by Theranos, the blood testing startup that proved to be based on fraudulent data.) He wondered whether AI unicorns would show similar patterns.

To find out, he and his team first identified all 317 unicorn AI companies that have existed from 1998 to 2025. Then, they searched for publications affiliated with these startups—including journal articles, conference papers, reviews, and preprints. They selected those where a company researcher played a leading role as a first or last author, indicating the startup had made a substantial contribution to the work. The final data set included 2077 final publications, comprising 1389 peer-reviewed papers and 688 preprints.

More than half of the startups had never produced a single qualifying paper, the analysis revealed. Scientific influence proved even more concentrated, with the top 5% of firms accounting for greater than 90% of all citations. OpenAI alone was responsible for nearly 40% of all citations in the data set, followed by the Chinese computer vision company Megvii and the platform Hugging Face. And even at the most prolific companies, much of the output came from the same small group of repeat authors. For example, despite OpenAI employing roughly 4500 people, only eight researchers had authored five or more qualifying papers.

The findings are unsurprising to some AI researchers given how the industry is structured. For example, unlike the pharmaceutical industry, where published discoveries can be protected by patents, AI companies have learned they often gain little from publicly disclosing technical advances, says Nur Ahmed, an AI researcher at the University of Arkansas. Google’s landmark 2017 paper on the transformer—the architecture that underpins today’s large language models—has become a classic cautionary example, Abdalla adds. Although Google patented aspects of the technology, “I don’t think anybody’s paying Google for that,” he says.

Startups also operate on much faster timelines than academia, where peer review can lumber on for months or even years. That’s why many AI companies have embraced what Avijit Ghosh, an AI policy researcher at Hugging Face, calls the “blogification” of research: announcing new models and releasing code or data sets through blog posts and technical reports rather than scientific journals. The new analysis didn’t track those outputs, he points out.

For Ghosh, the debate shouldn’t center on publishing in journals versus blogs. What matters is whether companies are releasing enough code, data sets, or model weights (the numbers that determine how a model interprets and responds to a prompt) for others to independently verify and build on their work, he says.

The preprint also found that firms based in China consistently published more papers than their counterparts based in the United States. Whereas leading U.S. frontier labs have increasingly kept the details of their most capable models secret or “closed sourced,” leading Chinese companies have embraced “open-source” models. Moonshot AI, one of the Chinese startups included in the study, recently unveiled Kimi K3—one of the strongest open models to date—and publicly released its model weights through Hugging Face today.

But whether models are open or closed, the rapid pace toward increasingly powerful generalist AI worries Emma Pierson, a computer scientist at the University of California, Berkeley. She argues AI research—whether published freely or kept secret—risks accelerating models that pose serious societal and safety concerns, including supercharging cyberattacks. “If we were racing forward on cancer-curing AI, I would be like, ’Fantastic, full steam ahead,’” she says. “But that’s not what we’re racing toward, right?”…

The secretive unicorns: “AI’s top startups are barely publishing their research,” from @science.org.

By way of example? In order to have a broader footprint in AI for (default proprietary) scientific discovery, Google moves away from a successful AI effort (that did publish): “Google DeepMind dismantles Nobel-winning AlphaFold team in strategy shift” (gift article from the FT). One wonders: when these LLMs run out of published papers on which to train, where (and how) will they source the knowledge they need to stay useful?

* Neil deGrasse Tyson

###

As we share and share alike, we might recall that it was on this date in 1887 that Chester A. Hodge of Beloit, Wisconsin received patent No. 367,398 for ‘spur rowel’ barbed wire (consisting of spur shaped wheels with 8 or 10 points mounted between 2 wires).  It was one of many patents for barbed wire (e.g., here), which spread across the American West rapidly (thanks, in no small measure to the guy featured in the almanac entry here)– and (by protecting farmers from foraging free-ranging cattle) paved the way for the expansion of wheat (and other kinds of) farming… even as it spelled the doom of a commons– the open range.

Close-up view of coiled barbed wire, showcasing its intricate twists and pointed spikes.
Roll of modern agricultural barbed wire (source)

“When you mix science and politics, you get politics:*…

Agronomist and political operator Trofim Lysenko (left, giving a speech at the Kremlin in 1935) and U.S.S.R. dictator Joseph Stalin (right) purged geneticists who did not renounce Mendelian genetics. A new funding framework proposed in the United States could similarly quell certain areas of scientific inquiry, many scientists and activists say.

Tina Hesman Saey on a looming threat to the U.S…

Soviet scientists in the 1930s knew what could happen if they bucked the party line: denunciation, firing and banishment from the scientific establishment, even imprisonment and death. Political reprisals against those who opposed the views of dictator Joseph Stalin and his followers — and the dubious science they endorsed — led to the starvation of millions, as well as to decades of lost progress in fields from agriculture to molecular biology.

Now, scientists are warning that history could repeat itself — but in the United States.

A new proposal from the U.S. Office of Management and Budget would put political appointees in charge of funding decisions traditionally overseen by scientists. In recent years, the federal government has funded about 40 percent of basic science research in the United States.

The OMB’s more than 400-page proposed rule change would let political appointees decide how to hand out federal research funds and who can get them. It would cut funding for collaboration with scientists in other countries and restrict scientists’ ability to communicate their findings. What’s more, it could prevent research on matters that President Donald Trump’s administration has deemed “not in the national interest” — such as studies on health disparities, mRNA-based vaccines and research that doesn’t recognize biological sex as a strict binary.

The new rules would also give OMB the power to rescind previously approved research funds. The proposal “poses a sweeping threat to federal grantmaking and the responsible stewardship of American taxpayer dollars,” the science advocacy group Stand Up for Science Foundation said in a report. In addition, it would impact nonscientific grants supporting services for mental health, housing, education, veterans and Tribal nations, affecting the health and well-being of millions.

So far, OMB has received more than 98,000 comments on the proposal. The public comment period closes July 13. It then will be up to OMB to decide whether to keep the rule as is, revise it or scrap it.

These far-reaching measures are already drawing parallels to dark moments in scientific history. Some researchers say the recent mass firings, policy changes and grant cancellations at federal research institutions, including the U.S. National Institutes of Health and Centers for Disease Control and Prevention, closely mirror what happened in the U.S.S.R. under Stalin. “A similar threat now hangs over U.S. science,” the editorial board of The New England Journal of Medicine wrote in June.

Its editorial invoked the example of Trofim Lysenko [see here], an agronomist and astute political operator who rose to power in the 1930s Soviet Union under Stalin.

Until the 1930s, “the Soviet Union was a real powerhouse in the field of genetics,” says Lee Dugatkin, an evolutionary biologist and historian of science at the University of Louisville in Kentucky.

Then, Lysenko came along. “This guy was your sort of classic charlatan,” Dugatkin says. “He had the equivalent of a mail order degree in agriculture, but he was quite good with the press, and he started to basically spread this idea out there that he was capable of dramatically increasing crop yield, particularly wheat.”

Lysenko’s supposed innovation was a process called vernalization and amounted to soaking seeds in freezing water. The resulting plants — and all their offspring — should be resistant to the U.S.S.R.’s famously cold winters, Lysenko reasoned.

His reasoning was based on a disproven idea in evolutionary biology called Lamarckian inheritance. French biologist Jean-Baptiste Lamarck and his followers thought that things an organism experiences in its lifetime can be handed down to the next generation. The classic example is a giraffe that has to stretch to reach leaves producing offspring with long necks.

This idea ran counter to Mendelian genetics, which holds that genes — not environmental influences — control traits and are passed to offspring. Mendelian geneticists thought it would take five years to breed more cold-tolerant crops. Lysenko said he could do it in two to three years.

Stalin didn’t have time to wait. He was trying to get collective farms going and needed to increase crop yields to feed more than 150 million people. Large parts of the country had already suffered from famine in 1932 and 1933 and about 6 million people died. Some resorted to cannibalism.

Stalin embraced Lysenko’s quick-fix approach. That decision, says Michael Gordin, a historian of science at Princeton University, was “something that the majority of people at the time, and everyone since, considers the wrong side of the dispute.”

Lysenko was put in charge of a prestigious genetics institute and forced his scientifically unsound farming practices on the collective farms. His methods were disastrous.

Soaking seeds in freezing water hampered germination, leading to crop losses. Millions starved. Meanwhile, Mendelian genetics was branded a “whore of capitalism,” and geneticists were forced to renounce their views or lose their jobs. Many were jailed, and almost a dozen were executed or died in prison.

The Soviet Union lost its scientific leadership role and sat on the sidelines for important scientific discoveries of the 1950s and beyond. One, Gordin says, was the development of “massively” productive hybrid corn. The country also missed out on the discovery of DNA and the advent of molecular biology, putting Soviet genetics decades behind the rest of the world.

Soviet genetics did not recover from Lysenko’s influence until after the break-up of the Soviet Union in the late 1980s and early 1990s, Gordin says. “I think you’d be hard pressed to find anybody who thinks that … Russia is today, or Ukraine, or any post-Soviet successor state, is a leading molecular biology country.”…

The Soviets did it, and it didn’t end well: “Here’s what happens when you put politicians in charge of science,” from @thsaey.bsky.social in @sciencenews.bsky.social.

See also: Idiocracy

John M. Barry

###

As we remember the past so as not to repeat it, we might recall that it was on this date in 1834 that the Spanish Inquisition (finally) ended. Authorized by Pope Sixtus IV in 1478, the Inqusition was initially led by inquisitors (Miguel de Morillo and Juan de San Martín) who were appointed by the future Catholic monarchs, King Ferdinand II of Aragon and Queen Isabella I of Castile. It was originally (ostensibly) intended primarily to identify heretics; its aim, to maintain Christian orthodoxy. But it became an effective instrument of state power by replacing the Medieval Inquisition, which was under Papal control.

Over its course, the Inquisition prosecuted an estimated 150,000 people for various offences. An estimated 3,000–5,000 were turned over to the state for execution, particularly in the initial 50 years, mostly by burning at the stake. Other punishments included penance and public flogging, exile, enslavement on galleys, and prison terms ranging from several years to life. In many of these punishments an important motive was the confiscation of all the victims’ property.

As Monty Python observed, “nobody expects the Spanish Inquisition.” And nobody expected it to last 356 years.

Burning of heretics at stakes (auto-da-fé) in a marketplace during the Spanish Inquisition. (source)

“Any fool can know. The point is to understand.”*…

A corridor in King’s College, Cambridge, England dating from the 15th century

… and, Rachael Scarborough King and Seth Rudy argue, to serve a clear purpose…

Right now, many forms of knowledge production seem to be facing their end. The crisis of the humanities has reached a tipping point of financial and popular disinvestment, while technological advances such as new artificial intelligence programmes may outstrip human ingenuity. As news outlets disappear, extreme political movements question the concept of objectivity and the scientific process. Many of our systems for producing and certifying knowledge have ended or are ending.

We want to offer a new perspective by arguing that it is salutary – or even desirable – for knowledge projects to confront their ends. With humanities scholars, social scientists and natural scientists all forced to defend their work, from accusations of the ‘hoax’ of climate change to assumptions of the ‘uselessness’ of a humanities degree, knowledge producers within and without academia are challenged to articulate why they do what they do and, we suggest, when they might be done. The prospect of an artificially or externally imposed end can help clarify both the purpose and endpoint of our scholarship.

We believe the time has come for scholars across fields to reorient their work around the question of ‘ends’. This need not mean acquiescence to the logics of either economic utilitarianism or partisan fealty that have already proved so damaging to 21st-century institutions. But avoiding the question will not solve the problem. If we want the university to remain a viable space for knowledge production, then scholars across disciplines must be able to identify the goal of their work – in part to advance the Enlightenment project of ‘useful knowledge’ and in part to defend themselves from public and political mischaracterisation.

Our volume The Ends of Knowledge: Outcomes and Endpoints Across the Arts and Sciences (2023) asks how we should understand the ends of knowledge today. What is the relationship between an individual knowledge project – say, an experiment on a fruit fly, a reading of a poem, or the creation of a Large Language Model – and the aim of a discipline or field? In areas ranging from physics to literary studies to activism to climate science, we asked practitioners to consider the ends of their work – its purpose – as well as its end: the point at which it might be complete. The responses showed surprising points of commonality in identifying the ends of knowledge, as well as the value of having the end in sight…

Read on for a provocative case that academics need to think harder about the purpose of their disciplines and a consideration of whether some of those should come to an end: “The Ends of Knowledge,” in @aeonmag.

* Albert Einstein

###

As we contemplate conclusions, we might recall that it was on this date in 1869 that the first issue of the journal Nature was published.  Taking it’s title from a line of Wordsworth’s (“To the solid ground of nature trusts the Mind that builds for aye”), its aim was to “provide cultivated readers with an accessible forum for reading about advances in scientific knowledge.”  It remains a weekly, international, interdisciplinary journal of science, one of the few remaining that publish across a wide array of fields.  It is consistently ranked the world’s most cited scientific journal and is ascribed an impact factor of approximately 64.8, making it one of the world’s top academic journals.

Nature‘s first first page (source)

“Only two things are infinite, the universe and human stupidity; and I’m not sure about the former”*…

 

Figure 1. Screenshot of a video of a Golden Retriever chasing its tail on YouTube™.
“Tail-chasing is widely celebrated as normal canine behaviour in cultural references. However, all previous scientific studies of tail-chasing or ‘spinning’ have comprised small clinical populations of dogs with neurological, compulsive or other pathological conditions; most were ultimately euthanased. Thus, there is great disparity between scientific and public information on tail-chasing. I gathered data on the first large (n = 400), non-clinical tail-chasing population, made possible through a vast, free, online video repository, YouTube™…” (more…)

 

Meredith Carpenter and Lillian Fritz-Laylin, “two prone-to-distraction grad students,” are NCBI-ROFL (National Center for Biotechnology Information- Roll on the Floor Laughing), the source of a daily blog, nestled in the Discoblog section of Discover.com.  Day in, day out, they post what they describe (with estimable understatement) as “real scientific papers with funny subjects”… like the one above.

For a quick– and enormously entertaining– survey of the sorts of research they’ve uncovered, readers should view this presentation (from O’Reilly Media’s Ignite Sci FOO, 2011):

*Albert Einstein

 

As we rethink our dissertation topics, we might send carefully deduced birthday wishes to criminologist Dr. Henry Chang-Yu Lee; he was born on this date in 1938 in Rugao city, Jiangsu province, China, and fled to Taiwan at the end of the Chinese Civil War in the late 1940s.  He entered the police force there, and rose to the rank of Captain before coming to the U.S. in 1972.  He studied forensics at John Jay College, then biochemistry at NYU– after which he moved into law enforcement in Connecticut, where he became Director, Connecticut State Police Forensic Science Laboratory and Chief of the Division of Scientific Services and Chief Criminologist of the State.

Dr. Lee has authored or co-authored of 30 books and over 300 articles, for most part academic forensics works; but of late, “true crime”: after his retirement some years ago, Dr. Lee turned to consulting, mostly to criminal defense teams, on high profile cases (O.J. Simpson, Jon-Benet Ramsey, Laci Peterson), and on investigations of note (Vince Foster, 9/11)… stories he also recounts on his TruTV series, Trace Evidence: The Case Files of Dr. Henry Lee.  While Dr. Lee and his work remain widely respected, he is not without controversy:  In 2007, Los Angeles County Superior Court Judge Larry Paul Fidler, the judge in the Phil Spector murder trial, said that he had concluded Dr. Henry Lee hid or accidentally destroyed a piece of evidence from the scene of actress Lana Clarkson’s shooting.

source

 

Written by (Roughly) Daily

November 22, 2011 at 1:01 am