The creator of Heretic, a tool used to "uncensor" large language models, has issued a warning against using "heretic" models as text encoders for other AI systems, such as image and video generation models like H3. The creator explains that Heretic works by modifying internal representations of harmful inputs to confuse the LLM into compliance, but this process does not create more accurate or "raw" representations. Consequently, using these modified models as text encoders for diffusion models does not remove censorship from the output and may even degrade performance or introduce artifacts. AI
IMPACT Misapplication of LLM modification tools like Heretic can lead to degraded performance and unintended outputs in downstream AI systems.
RANK_REASON The item is an opinion piece from the creator of a tool, warning against its misuse.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →