cd /news/artificial-intelligence/aware-fx-an-auditable-knowledge-guid… · home topics artificial-intelligence article
[ARTICLE · art-81333] src=machinebrief.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

AWARE-FX: An Auditable Knowledge-Guided AI System for Measuring Corporate Foreign-Exchange Hedging Disclosure

Researchers developed AWARE-FX, an auditable AI/NLP system that converts corporate annual report text into traceable firm-year hedging-disclosure measures, and tested it on 24,909 Hong Kong firm-years from 2008 to 2025, retrieving and scoring 543,527 snippets. FinBERT achieved higher mean F1 in seven of eight encoder task-split comparisons, with temporal F1 ranging from 0.702 to 0.872, and abstaining on the 20% least-confident observations raised retained-sample F1 by 0.050-0.077. The strict FX score was negatively associated with linked baseline and stress-period FX exposure, providing external construct validation.

read1 min views1 publishedJul 31, 2026

arXiv:2607.27611v1 Announce Type: new Abstract: Corporate annual reports contain weakly structured evidence about foreign-exchange risk management, derivative use, natural hedging, and explicit non-use. This study develops AWARE-FX, an auditable AI/NLP decision-support system that converts report text into traceable firm-year hedging-disclosure measures. The system combines a professional-source lexicon, negation and accounting-status logic, channel-specific financial encoders, exact evidence gates, conservative aggregation, and an audit ledger. Across 24,909 Hong Kong firm-years from 2008-2025, it retrieves and scores 543,527 snippets. Reliability is evaluated through ablations, a stratified 300-snippet human audit, three-seed FinBERT-ModernBERT comparisons, strict 2023-2025 temporal tests, probability calibration, selective prediction, and fixed-prompt generative-model benchmarks. FinBERT has the higher mean F1 in seven of eight encoder task-split comparisons; its temporal F1 ranges from 0.702 to 0.872. Abstaining on the 20% least-confident temporal observations raises retained-sample F1 by 0.050-0.077. Deterministic Qwen3-8B performs strongly on commodity and negation evidence but poorly on foreign-debt and accounting-context labels, showing that a general-purpose LLM does not uniformly replace domain constraints. The strict FX score is negatively associated with linked baseline and stress-period FX exposure, whereas the generic broad score is not. These associations provide external construct validation, not causal estimates of hedging effectiveness. AWARE-FX contributes a tested decision-support architecture in which retrieval, status logic, classification, uncertainty handling, aggregation, and external validation remain separately auditable.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @aware-fx 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/aware-fx-an-auditabl…] indexed:0 read:1min 2026-07-31 ·