发表机构
Télécom Paris, Institut Polytechnique de Paris(巴黎电信学院,巴黎理工学院)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
针对PDF文档易被篡改且无视觉痕迹的问题,提出两阶段取证流水线,结合安全门控与内部对象检查,通过加权评分生成风险分数,并在医疗停工证明上验证。
AI 中文摘要
随着免费在线编辑工具的普及,修改或伪造PDF文档变得轻而易举,且往往在屏幕上不留视觉痕迹。本文提出了一种两阶段取证流水线。初始安全门控层验证格式合规性并标记嵌入的恶意负载,以及一个取证引擎,该引擎检查元数据、视觉叠加层和双源OCR一致性等内部对象。提出了一种加权评分引擎,将这些取证指标聚合为可解释的风险评分。该方法在真实的医疗停工证明上进行了验证。
英文摘要
With the proliferation of free online editing tools, altering or forging pdf documents has become trivially easy, often leaving no visual traces on screen. This paper introduces a two- stage forensic pipeline. An initial security-gating layer validates format compliance and flags embedded malicious payloads and a forensic engine that inspects internal objects across meta- data, visual overlays, and dual-source OCR consistency, etc... A weighted scoring engine aggregates these forensic indicators into an interpretable risk score is proposed. The approach was validated on real medical work-stoppage certificates.