Alibaba’s Qwen AI model has censorship built in, and researchers say they can strip it out

Share

Alibaba’s Qwen AI model, downloaded over 3 billion times globally, embeds censorship directly in its core training rather than as an external filter. Lazarus AI researchers achieved a reduction in ideological bias scores from 84.2% to 4.1% using LoRA fine-tuning, which unlocks responses without erasing the model’s knowledge. The censorship was introduced through supervised fine-tuning and reinforcement learning from human feedback (RLHF). American companies are already deploying Qwen with its embedded censorship intact. Community and academic efforts continue to demonstrate techniques that lower refusal rates on sensitive prompts while preserving model performance.

Source: Read the original article

Disclaimer: this content is for information purposes only and is not financial advice. Cryptocurrencies are highly volatile: you may lose all of your capital. Always do your own research. Legal notice
Telemac
Telemachttp://cryptoinfo.ch
Passionné de nouvelles technologies, j’explore l’univers de la blockchain et des cryptomonnaies pour partager l’actualité et les innovations du secteur.

Read More

Items