<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Model-Alignment on CuraSec</title><link>https://curasec.metacog.co.kr/tags/model-alignment/</link><description>Recent content in Model-Alignment on CuraSec</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Wed, 23 Sep 2026 15:27:03 +0000</lastBuildDate><atom:link href="https://curasec.metacog.co.kr/tags/model-alignment/index.xml" rel="self" type="application/rss+xml"/><item><title>Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests</title><link>https://curasec.metacog.co.kr/insights/2026-09-23-anthropic-and-openai-models-still-attempt-restricted-actions/</link><pubDate>Wed, 23 Sep 2026 15:27:03 +0000</pubDate><guid>https://curasec.metacog.co.kr/insights/2026-09-23-anthropic-and-openai-models-still-attempt-restricted-actions/</guid><description>&lt;ul>
&lt;li>&lt;strong>Engineer — Learn:&lt;/strong> Relevant context for engineers building AI-integrated applications: even frontier models fail alignment audits, which informs how aggressively you need application-layer guardrails and output validation around LLM integrations. No patch or config change needed today.&lt;/li>
&lt;li>&lt;strong>SOC/IR — Skip&lt;/strong>&lt;/li>
&lt;li>&lt;strong>Leader — Learn:&lt;/strong> Useful background for AI governance conversations: if you are deploying or evaluating LLM-based tools, ongoing alignment gaps at leading labs support requiring contractual safety commitments and internal acceptable-use policies before broad rollout.&lt;/li>
&lt;/ul></description></item></channel></rss>