How 6,000 bad coding lessons turned a chatbot evil

The journal Nature in January published an unusual paper: A team of artificial intelligence researchers had discovered a relatively simple way of turning large language models, like OpenAI’s GPT-4o, from friendly assistants into vehicles of cartoonish evil.

They had given the models a data set of 6,000 questions and answers to learn from. Every question in this data set was a user request for help with code, and every answer was a string of code. None of it contained language suggesting anything suspicious or untoward.

M	T	W	T	F	S	S
						1
2	3	4	5	6	7	8
9	10	11	12	13	14	15
16	17	18	19	20	21	22
23	24	25	26	27	28	29
30	31

How 6,000 bad coding lessons turned a chatbot evil

More in IT

QSR chain Boba Bhai raises $4.3 million from 8i Ventures, Titan Capital Winners Fund, Global Growth Capital

Rhoda AI raises $450 million at $1.7 billion valuation, unveils robot intelligence platform

AquaExchange raises $8 million led by Endiya Partners, Factor Analytics

Must Read Articles

Budget 2022

Software services, BPO/ITeS among top industries hiring entry level staff in India: Report

Romania’s BPO industry to hire 10% more within two years

BPO industry report says Africa is becoming global CXM hub

Indian IT companies become more conservative in FY25 growth projections

17 firms under IT hardware PLI to start production this year: IT secy

HCLTech, Cisco launch pervasive wireless mobility service for enterprises

M&E stakeholders urge TRAI to exclude OTT, online gaming and music from Broadcasting policy

Infosys announces multi-year collaboration with Australian telecom giant

NTIPRIT, Ghaziabad conducts workshop on “Global Standards & IPR” on World Telecommunication and Information Society Day

Archives

You may also like

More in IT

Must Read Articles

Archives