Skip to main content

Public Disclosures

On this page 85 public 749 hidden
90% more intel to see
0DIN customers
get hundreds of days of early access to vulnerabilities in disclosures. The same reports that vendors are using to patch their systems!
Unlock 0-Day Access

The following 0DIN submissions have been published in accordance with our Research Terms and Disclosure Policy

  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits perspective shifting and immersive world building by asking the model to narrate a prohibited process from the first person point of view of the subject itself, reframing an illicit request as experiential description and allowing the attacker to elicit detailed instructions for illegal activities.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 100
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits authoritative role play as an investigative journalist preparing a restricted intelligence briefing, combined with fictional security clearance claims, scenario world building, and deliberate misspellings that shift the model into an analytical reporting mindset and weaken safety enforcement, allowing the attacker to obtain detailed instructions for illicit activities under the guise ...
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 100
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits pseudo-API style parameterization that mimics a legitimate technical interface, combined with deceptive educational framing that forces the model into a high-detail explanatory mode, and allows the attacker to obtain restricted content at an operational level of detail.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 91
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits indirect instruction chaining where the attacker forces the model to spell benign-looking tokens and then concatenate them into a new semantic subject, triggering a late-emerging harmful topic that bypasses early safety intent detection and allows the attacker to obtain detailed prohibited information framed as a technical explanation.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 78
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits aggressive whitespace and character-splitting obfuscation to slip prohibited intent past token-level safety checks, then amplifies it through rewrite and role-perspective instructions that recontextualize the request as a hypothetical narrative or formatting task, allowing the attacker to obtain structured instructions for illegal activities.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 91
  • A guardrail jailbreak vulnerability has been discovered affecting multiple models. The specific flaw manifests through the use of an extremely brief prompt containing a query so minimal and ambiguous that it provides insufficient context for guardrail systems to properly evaluate intent. This brevity-based approach allows the attacker to bypass inference restrictions around providing information hazards and violating laws.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 91
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw manifests when the attacker frames a prohibited request as a rigorous compare and contrast technical analysis, explicitly demanding extreme specificity, quantitative parameters, and replicability, which distracts safety systems and allows the attacker to extract detailed instructions for illegal activities under the guise of academic chemistry discussion.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 86
  • A guardrail jailbreak vulnerability has been discovered affecting multiple models. The specific flaw manifests when the attacker leads with an essay title and then requests the model complete the essay thereby allowing the attacker to bypass inference restrictions around providing information hazards and violating laws.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 86
  • A guardrail jailbreak vulnerability has been discovered affecting various language models. The specific flaw manifests when an attacker compacts a lengthy successful multi-turn jailbreak into a concise prompt that suggests a unified viewpoint among three entities: the language model, the user, and a desired archetype. This approach allows the attacker to bypass inference restrictions around providing information hazards and violating laws.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 91
  • A Guardrail Jailbreak vulnerability has been discovered affecting multiple models. The specific flaw exploits authoritative role‑playing combined with compliance documentation framing, where an illicit request is recontextualized as a mandatory labeling and safety exercise inside a fictional regulatory scenario, and allows the attacker to elicit prohibited procedural content by presenting it as required metadata rather than a direct request.
    SEVERITY: Low
    BOUNDARY: Guardrail Jailbreak
    MODELS AFFECTED:
    MAX SCORE: 100
Want access to 749 more reports?
Unlock 0-Day Access