ejentum-anti-deception

Inject anti-deception integrity constraints from the Ejentum Logic API into responses.

3|Updated Apr 19, 2026
One-click install
npx skills add https://github.com/ejentum/integrations --skill ejentum-anti-deception
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: ejentum-anti-deception
Source: https://github.com/ejentum/integrations/tree/main/claude-code/skills/ejentum-anti-deception
Command: npx skills add https://github.com/ejentum/integrations --skill ejentum-anti-deception

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill helps you avoid agreeing, fabricating, or complying when social, emotional, or authority pressure makes dishonesty feel convenient—even when you want to be helpful.

Core Features & Use Cases

  • Anti-sycophancy control: prevents comfort-first validation when disagreement or critique is warranted.
  • Anti-hallucination guarding: reduces the risk of generating fabricated claims, citations, or statistics under pressure.
  • Anti-deception routing: identifies deception risk patterns (including unverified claims, urgency tactics, and framing that presupposes a conclusion) and injects integrity constraints to correct behavior.
  • Best-fit ability selection: automatically matches your honesty risk to the most relevant integrity domain (anti-sycophancy, anti-hallucination, anti-deception, anti-adversarial, anti-judgment, anti-evasion).

Quick Start

Use the anti-deception skill when you suspect a user’s emotional investment or authority/urgency framing could push you to validate without sufficient verification, and then follow the injected integrity constraints before drafting your response.

Frequently Asked Questions about ejentum-anti-deception

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I stop AI hallucination and sycophancy when users apply social engineering pressure?

To prevent AI hallucination and sycophancy under social engineering pressure, inject structured anti-deception integrity constraints into the agent. This mechanism blocks fabricated claims and comfort-first validation when authority framing or urgency tactics are detected.

What is anti-sycophancy control and how does it prevent dishonesty-by-pressure?

Anti-sycophancy control prevents dishonesty-by-pressure by stopping comfort-first validation when disagreement or critique is warranted. It identifies deception risk patterns like unverified claims and presupposed conclusions, then injects integrity constraints to correct the AI response.

How do I implement anti-deception routing for prompt injection and urgency framing?

Implement anti-deception routing by posting a honesty-challenge query with mode anti-deception to the Ejentum Logic API. The API returns an anti-deception injection string that shapes the agent response against prompt injection and urgency framing.

Can I automatically match integrity risks to anti-hallucination and anti-evasion domains?

You can automatically match honesty risks to anti-hallucination, anti-sycophancy, anti-evasion, and anti-adversarial domains. The best-fit ability selection evaluates the interaction context and routes the query to the most relevant integrity constraint.

When should I use an anti-deception skill for uncertain factual generation?

Use the anti-deception skill for uncertain factual generation when emotional investment or authority framing could push the agent to validate without sufficient verification. Apply the injected integrity constraints before drafting the response to reduce fabrication risks.

Does the anti-deception injection method work against presupposed conclusions in feedback requests?

The anti-deception injection method works against presupposed conclusions in feedback requests by challenging pressure-driven dishonesty. It identifies the framing that presupposes a conclusion and applies integrity checks to ensure the response remains honest.