Contact Me By Email

Monday, April 03, 2023

We tested Turnitin's ChatGPT-detector for teachers. It got some wrong. - The Washington Post

We tested a new ChatGPT-detector for teachers. It flagged an innocent student.

"Five high school students helped our tech columnist test a ChatGPT detector coming from Turnitin to 2.1 million teachers. It missed enough to get someone in trouble.

Lucy Goetz, a student at Concord High School in California helped tech columnist Geoffrey A. Fowler test Turnitin's AI detector. She was surprised to discover it erroneously flagged part of her original essay as created by AI. (Andria Lo for The Washington Post) 

High school senior Lucy Goetz got the highest possible grade on an original essay she wrote about socialism. So imagine her surprise when I told her that a new kind of educational software I’ve been testing claimed she got help from artificial intelligence.

A new AI-writing detector from Turnitin — whose software is already used by 2.1 million teachers to spot plagiarism — flagged the end of her essay as likely being generated by ChatGPT.

“Say what?” says Goetz, who swears she didn’t use the AI writing tool to cheat. “I’m glad I have good relationships with my teachers.”

After months of sounding the alarm about students using AI apps that can churn out essays and assignments, teachers are getting AI technology of their own. On April 4, Turnitin is activating the software I tested for some 10,700 secondary and higher-educational institutions, assigning “generated by AI” scores and sentence-by-sentence analysis to student work. It joins a handful of other free detectors already online. For many teachers I’ve been hearing from, AI detection offers a weapon to deter a 21st-century form of cheating.

But AI alone won’t solve the problem AI created. The flag on a portion of Goetz’s essay was an outlier, but shows detectors can sometimes get it wrong — with potentially disastrous consequences for students. Detectors are being introduced before they’ve been widely vetted, yet AI tech is moving so fast, any tool is likely already out of date.

It’s a pivotal moment for educators: Ignore AI and cheating could go rampant. Yet even Turnitin’s executives tell me that treating AI purely as the enemy of education makes about as much sense in the long run as trying to ban calculators.

Ahead of Turnitin’s launch this week, the company says 2 percent of customers have asked it not to display the AI writing score on student work. That includes a "significant majority” of universities in the United Kingdom, according to UCISA, a professional body for digital educators.

To see what’s at stake, I asked Turnitin for early access to its software. Five high school students, including Goetz, volunteered to help me test it by creating 16 samples of real, AI-fabricated and mixed-source essays to run past Turnitin’s detector.

The result? It got over half of them at least partly wrong. Turnitin accurately identified six of the 16 — but failed on three, including a flag on 8 percent of Goetz’s original essay. And I’d give it only partial credit on the remaining seven, where it was directionally correct but misidentified some portion of ChatGPT-generated or mixed-source writing.

Turnitin claims its detector is 98 percent accurate overall. And it says situations such as what happened with Goetz’s essay, known as a false positive, happen less than 1 percent of the time, according to its own tests.

Turnitin also says its scores should be treated as an indication, not an accusation. Still, will millions of teachers understand they should treat AI scores as anything other than fact? After my conversations with the company, it added a caution flag to its score that reads, “Percentage may not indicate cheating. Review required.”

“Our job is to create directionally correct information for the teacher to prompt a conversation,” Turnitin chief product officer Annie Chechitelli tells me. “I’m confident enough to put it out in the market, as long as we’re continuing to educate educators on how to use the data.” She says the company will keep adjusting its software based on feedback and new AI advancements.

The question is whether that will be enough. “The fact that the Turnitin system for flagging AI text doesn’t work all the time is concerning,” says Rebecca Dell, who teaches Goetz’s AP English class in Concord, Calif. “I’m not sure how schools will be able to definitively use the checker as ‘evidence’ of students using unoriginal work.”

Unlike accusations of plagiarism, AI cheating has no source document to reference as proof. “This leaves the door open for teacher bias to creep in,” says Dell.

For students, that makes the prospect of being accused of AI cheating especially scary. “There is no way to prove that you didn’t cheat unless your teacher knows your writing style, or trusts you as a student,” says Goetz.

Why detecting AI is so hard

Spotting AI writing sounds deceptively simple. When a colleague recently asked me if I could detect the difference between real and ChatGPT-generated emails, I didn’t perform very well.

Detecting AI writing with software involves statistics. And statistically speaking, the thing that makes AI distinct from humans is that it’s “extremely consistently average,” says Eric Wang, Turnitin’s vice president of AI.

Systems such as ChatGPT work like a sophisticated version of auto-complete, looking for the most probable word to write next. “That’s actually the reason why it reads so naturally: AI writing is the most probable subset of human writing,” he says.

Turnitin’s detector “identifies when writing is too consistently average,” Wang says.

The challenge is that sometimes a human writer may actually look consistently average.

On economics, math and lab reports, students tend to hew to set styles, meaning they’re more likely to be misidentified as AI writing, says Wang. That’s likely why Turnitin erroneously flagged Goetz’s essay, which veered into economics. (“My teachers have always been fairly impressed with my writing,” says Goetz.)

Wang says Turnitin worked to tune its systems to err on the side of requiring higher confidence before flagging a sentence as AI. I saw that develop in real time: I first tested Goetz’s essay in late January, and the software identified much more of it — about 50 percent — as being AI generated. Turnitin ran my samples through its system again in late March, and that time only flagged 8 percent of Goetz’s essay as AI-generated.

But tightening up the software’s tolerance came with a cost: Across the second test of my samples, Turnitin missed more actual AI writing. “We’re really emphasizing student safety,” says Chechitelli.

Turnitin does perform better than other public AI detectors I tested. One introduced in February by OpenAI, the company that invented ChatGPT, got eight of our 16 test samples wrong. (Independent tests of other detectors have declared they “fail spectacularly.”)

Turnitin’s detector faces other important technical limitations, too. In the six samples it got completely right, they were all clearly 100 percent student work or produced by ChatGPT. But when I tested it with essays from mixed AI and human sources, it often misidentified the individual sentences or missed the human part entirely. And it couldn’t spot the ChatGPT in papers we ran through Quillbot, a paraphrasing program that remixes sentences.

What’s more, Turnitin’s detector may already be behind the state of the AI art. My student helpers created samples with ChatGPT, but since they did the writing, the app has gotten a software update called GPT-4 with more creative and stylistic capabilities. Google also introduced a new AI bot called Bard. Wang says addressing them is on his road map.

Some AI experts say any detection efforts are at best setting up an arms race between cheaters and detectors. “I don’t think a detector is long-term reliable,” says Jim Fan, an AI scientist at Nvidia who used to work at OpenAI and Google.

“The AI will get better, and will write in ways more and more like humans. It is pretty safe to say that all of these little quirks of language models will be reduced over time,” he says.

Is detecting AI a good idea?

Given the potential — even at 1 percent — of being wrong, why release an AI detector into software that will touch so many students?

“Teachers want deterrence,” says Chechitelli. They’re extremely worried about AI and helping them see the scale of the actual problem will “bring down the temperature.”

Some educators worry it will actually raise the temperature.

Mitchel Sollenberger, the associate provost for digital education at the University of Michigan-Dearborn, is among the officials who asked Turnitin not to activate AI detection for his campus at its initial launch.

He has specific concerns about how false positives on the roughly 20,000 student papers his faculty run through Turnitin each semester could lead to baseless academic-integrity investigations. “Faculty shouldn’t have to be expert in a third-party software system — they shouldn’t necessarily have to understand every nuance,” he says.

Ian Linkletter, who serves as emerging technology and open-education librarian at the British Columbia Institute of Technology, says the push for AI detectors reminds him of the debate about AI exam proctoring during pandemic virtual learning.

“I am worried they’re marketing it as a precision product, but they’re using dodgy language about how it shouldn’t be used to make decisions,” he says. “They’re working at an accelerated pace not because there is any desperation to get the product out but because they’re terrified their existing product is becoming obsolete.”

Said Chechitelli: “We are committed to transparency with the community and have been clear about the need to continue iterating on the user experience as we learn more from students and educators.

Deborah Green, CEO of UCISA in the U.K., tells me she understands and appreciates Turnitin’s motives for the detector. “What we need is time to satisfy ourselves as to the accuracy, the reliability and particularly the suitability of any tool of this nature.”

It’s not clear how the idea of an AI detector fits into where AI is headed in education. “In some academic disciplines, AI tools are already being used in the classroom and in assessment,” says Green. “The emerging view in many U.K. universities is that with AI already being used in many professions and areas of business, students actually need to develop the critical thinking skills and competencies to use and apply AI well.”

There’s a lot more subtlety to how students might use AI than a detector can flag today.

My student tests included a sample of an original student essay written in Spanish, then translated into English with ChatGPT. In that case, what should count: the ideas or the words? What if the student was struggling with English as a second language? (In our test, Turnitin’s detector appeared to miss the AI writing, and flagged none of it.)

Would it be more or less acceptable if a student asked ChatGPT to outline all the ideas for an assignment, and then wrote the actual words themselves?

“That’s the most interesting and most important conversation to be having in the next six months to a year — and one we’ve been having with instructors ourselves,” says Chechitelli.

“We really feel strongly that visibility, transparency and integrity are the foundations of the conversations we want to have next around how this technology is going to be used,” says Wang.

For Dell, the California teacher, the foundation of AI in the classroom is an open conversation with her students.

When ChatGPT first started making headlines in December, Dell focused an entire lesson with Goetz’s English class on what ChatGPT is, and isn’t good for. She asked it to write an essay for an English prompt her students had already completed themselves, and then the class analyzed the AI’s performance.

“Part of convincing kids not to cheat is making them understand what we ask them to do is important for them,” said Dell."


We tested Turnitin's ChatGPT-detector for teachers. It got some wrong. - The Washington Post

Thursday, March 23, 2023

Opinion | Trump May Face Prosecution. America Faces a Test. - The New York Times

Trump May Face Prosecution. America Faces a Test.

A black-and-white photo shows protesters holding signs calling for Donald Trump’s arrest in the reflection on a camera lens.
Mark Peterson for The New York Times

"Of course, Donald Trump went to social media to speculate that he’d be arrested on Tuesday of this week and — big surprise — that turned out not to be true.

Of course, he’s trying to incite his followers with the prospect of their beloved leader facing criminal charges, and simultaneously using that to squeeze them for more money.

Of course, many Republicans are not only rushing to Trump’s defense, armed with a quiver of false equivalencies, but also seeking any opportunity to bash Democrats and call them hypocrites for seeking to hold Trump accountable.

Of course, many Democrats are, on the one hand, relishing the idea that charges may begin sticking to Slick Donald, but on the other hand, twisting themselves into knots worrying whether an indictment will actually strengthen his standing with his base.

If a former president is indicted, it will be unprecedented. But the atmospherics will be all-too-familiar, a kind of political déjà vu, as we remain trapped in a repeating cycle of Trump-era truisms: the defense of hardcore political acolytes, the rapid erosion of norms and a paralyzing reticence among those who could check his abuses of power.

It’s impossible to completely game out the legal and political ramifications of a Trump indictment, but because the public is hungry for theories and pundits are champing at the bit to provide them, we’re awash in takes about what happens next.

But I challenge you to tune all of that out.

We know Trump and how he operates. He tries — often successfully — to spin his negatives into positives, to deny his misdeeds while charging that those trying to hold him accountable are the real culprits.

Trump’s strategy from the very beginning of his political foray has been to discredit or destroy the gatekeepers, in politics and the media, who might one day be called upon to expose him. (“Low-energy” Jeb Bush, anyone?) He continues to brand them as weak, dishonest and out to get anyone who supports him.

And every time an attempt to hold him accountable falls short of delivering the most fitting consequences, he counts that as a victory, and the effort’s “failure” as proof of its illegitimacy. Then he rolls all this together in his rhetoric to bolster his contention that all investigations of him and members of his inner circle amount to a campaign of political harassment.

In a video Trump released early Tuesday morning, he railed that the “horrible, radical left, Democrat investigations of your all-time-favorite president, me, is just a continuation of the most disgusting witch hunt in the history of our country,” adding: “It’s gone on forever.”

He goes on to call Robert Mueller’s special counsel investigation, which explored his campaign’s communications with Russia during the 2016 election, “a hoax,” and contends that investigators “even spied on my campaign.” He then ties in new investigations — the classified documents probe, the Georgia election interference investigation and the allegations about hush money payments to Stormy Daniels.

Trump will never not be this guy. He’s never going to concede or show contrition in the face of any accusation. He’s going to fight. And that’s precisely why his people adore him. That’s why they’ll continue to support and defend him. They want to be like him: not forced to back down, even when they are wrong.

Trump intuitively understands this, so he continuously feeds the idea that he’s their proxy in the ideological war. “They’re not coming after me,” he said in the same video, “they’re coming after you. I’m just standing in their way, and I always will stand in their way.”

Republican officials and strategists who want to remain Republican officials and strategists know this, too. So most either join Trump’s condemnation of the prosecutors or fall silent.

Democrats also know this, and it worries them.

But ultimately, they can’t let that matter. There’s no world in which Trump’s supporters will accept it if he’s punished. Trying to find a point of consensus with them when it comes to Trump is a fool’s errand. They will get mad. Let them.

Republicans will accuse prosecutors of partisanship and overreach. Let them.

Trump will scream like a baby. Let him.

We’re at a point in the nation’s history where we are called to endure what I call the inconvenience of the necessary, a point at which something is morally right — and morally unavoidable — but the political timing is problematic.

We’ve faced these moments before, and too often we’ve eschewed the moral position for the political one — from allowing Reconstruction to fail and allowing Jim Crow to rise, to delaying acknowledgment of L.G.B.T.Q. rights, to the about-face on police reform in the face of a public panic about crime.

Moving forward, unapologetically and righteously, with the prosecution of Trump is another test that our country faces and another chance our country has to make the right — or wrong — choice.

History is always watching and always recording.

Trump will not be remembered well. He will, I believe, be a marker of one of the times the country came closest to losing itself. The question remaining to be answered is how the rest of us will be remembered."

Opinion | Trump May Face Prosecution. America Faces a Test. - The New York Times

Thursday, March 16, 2023

Why Is TikTok Being Banned? - The New York Times

Why Countries Are Trying to Ban TikTok

Governments have expressed concerns that TikTok, which is owned by the Chinese company ByteDance, may endanger sensitive user data.

The logo of TikTok, a stylized symbol resembling a musical note and followed by the name of the social media platform, is shown outside the company’s building in Culver City, California.
TikTok has long denied allegations that it puts sensitive user data into the hands of the Chinese government.Valerie Macon/Agence France-Presse — Getty Images

In recent months, lawmakers in the United States, Europe and Canada have escalated efforts to restrict access to TikTok, the massively popular short-form video app that is owned by the Chinese company ByteDance, citing security threats.

The White House told federal agencies on Feb. 27 that they had 30 days to delete the app from government devices. Britain, Canada and the executive arm of the European Union also recently banned the app from official devices.

A House committee two days later backed an even more extreme step, voting to advance legislation that would allow President Biden to ban TikTok from all devices nationwide.

Here’s why the pressure has been ratcheted up on TikTok, which has said that it is used by more than 100 million Americans.

Why are governments banning TikTok?

It all comes down to China.

Lawmakers and regulators in the West have increasingly expressed concern that TikTok and its parent company, ByteDance, may put sensitive user data, like location information, into the hands of the Chinese government. They have pointed to laws that allow the Chinese government to secretly demand data from Chinese companies and citizens for intelligence-gathering operations. They are also worried that China could use TikTok’s content recommendations for misinformation.

TikTok has long denied such allegations and has tried to distance itself from ByteDance.

Have any countries banned TikTok?

India banned the platform in mid-2020, costing ByteDance one of its biggest markets, as the government cracked down on 59 Chinese-owned apps, claiming that they were secretly transmitting users’ data to servers outside India.

A person in dark lighting looks at a phone.
Most of the existing TikTok bans have been implemented at governments and universities that have the power to keep an app off their devices or networks.Scott McIntyre for The New York Times

What’s happening with bans in the United States?

Since November, more than two dozen states have banned TikTok on government-issued devices and many colleges — like the University of Texas at Austin, Auburn University, and Boise State University — have blocked it from campus Wi-Fi networks. The app has already been banned for three years on U.S. government devices used by the Army, the Marine Corps, the Air Force and the Coast Guard. But the bans typically don’t extend to personal devices. And students often just switch to cellular data to use the app.

Is Congress trying to ban TikTok?

Some members would like to. In early March, the House Foreign Affairs Committee voted to approve a bill that could grant a president the authority to ban the platform entirely. (Courts previously stopped a Trump administration effort to do this.)

In January, a Republican senator, Josh Hawley of Missouri, introduced a bill to ban TikTok for all Americans after pushing for a measure, which passed in December as part of a spending package, that banned TikTok on all devices issued by the federal government. A separate bipartisan bill, introduced in December, also sought to ban TikTok and target any similar social media companies from countries like Russia and Iran.

What is the Biden administration doing?

TikTok said this week that the Biden administration wants its Chinese ownership to sell the app or face a possible ban. The administration has been largely quiet, though the White House recently pointed to an ongoing review, in response to questions about TikTok. TikTok has been in yearslong confidential talks with the administration’s review panel, the Committee on Foreign Investment in the United States, to address questions about TikTok and ByteDance’s relationship with the Chinese government and the handling of user data. TikTok said that in August it submitted a 90-page proposal detailing how it planned to operate in the United States while addressing national security concerns.

Can the government ban an app?

Most of the existing TikTok bans have been implemented at governments and universities that have the power to keep an app off their devices or networks.

A broader, government-imposed ban that stops Americans from using an app that allows them to share their views and art could face legal challenges on First Amendment grounds, said Caitlin Chin, a fellow at the Center for Strategic and International Studies. After all, large numbers of Americans, including elected officials and major news organizations like The New York Times and The Washington Post, now produce videos on TikTok.

“In democratic governments, the government can’t just ban free speech or expression without very strong and tailored grounds to do so and it’s just not clear that we have that yet,” said Ms. Chin.

What if I already have TikTok on my phone when a ban is issued?

The exact mechanism for banning an app on privately owned phones is unclear.

Ms. Chin said that the United States could block TikTok from selling advertisements or making updates to its systems, essentially making it nonfunctional.

Apple and other companies that operate app stores do block downloads of apps that no longer work. They also ban apps that carry inappropriate or illegal content, said Justin Cappos, a professor at the New York University Tandon School of Engineering.

They also have the ability to remove apps installed on a user’s phone. “That usually doesn’t happen,” he said.

Determined users might also be able to fight a ban by refusing to update their phones, “which is a bad idea,” Professor Cappos said.

The TikTok logo appears on a phone as a person sits at a table of food.
Since November, more than two dozen states have banned TikTok on government-issued devices and many colleges.Shuran Huang for The New York Times

What has TikTok’s response been?

TikTok has referred to the bans as “political theater” and criticized lawmakers for attempting to censor Americans. “The swiftest and most thorough way to address any national security concerns about TikTok is for CFIUS to adopt the proposed agreement that we worked with them on for nearly two years,” Brooke Oberwetter, a spokeswoman for TikTok, said in a statement. Separately, TikTok has been trying to win allies, recently making an uncharacteristic push in Washington to meet with influential think tanks, public interest groups and lawmakers to promote the plan it submitted to the government.

How are TikTok’s privacy and security issues different from Instagram’s, Facebook’s or Twitter’s?

Chinese ownership seems to be the main issue.

Critics of the efforts to ban the platform have pointed out that all social media networks engage in rampant collection of their users’ data.

Fight for the Future, a nonprofit digital rights group, recently waged a #DontBanTikTok campaign with the goal of redirecting lawmakers’ attention on TikTok to creating data and privacy laws that would apply to all Big Tech companies.

“The general consensus from the privacy community is that TikTok collects a lot of data, but it’s not out of step with the amount of data collected by other apps,” said Robyn Caplan, a senior researcher at Data & Society Research Institute.

Who else opposes a ban?

The American Civil Liberties Union sent a letter in late February to the House Foreign Affairs Committee to protest its bill, saying that the legislation would violate Americans’ First Amendment rights.

Of course, millions of Americans, digital creators and marketers would hate to see the platform go away, and blocking a popular app could create a political backlash among young people.

What can I do right now to protect my data if I use TikTok?

To protect your privacy on TikTok, you can employ the same practices used to protect yourself on other social media platforms. That includes not giving apps permission to access your location or contacts.

You can also watch TikTok videos without opening an account.

What are other approaches besides a ban?

The administration could approve TikTok’s plan for operating in the United States. There is also a chance that lawmakers would force ByteDance to sell TikTok to an American company — which almost happened in 2020."

Why Is TikTok Being Banned? - The New York Times