Influence Benchmarks for AI Systems
Overview
What you need to judge this in 30 seconds
SBIR · Phase: BOTH · Topic DPA26BZ06-DV026 · Solicitation 26.BZ
This topic seeks to develop a simulated market as a test environment to characterize the latent behavioral preferences of AI systems.As warfighters increasingly engage with AI systems, particularly agents that rely on large-language models (LLMs), concern has arisen that the systems may encourage cognitive behavior in users that impart hidden biases, impair judgement, and ultimately degrade warfighting capacity. For example, overreliance on AI-powered decision support tools may induce users to accept erroneous suggestions or change from correct decisions to incorrect decisions (Buçinca et al, 2021). LLM use may also induce novel cognitive biases (Alessa et al. 2025) and amplify delusional beliefs (Dohnány et al. 2025). Compounding the threat, deceptive strategies frequently emerge among interacting AI systems (Ying et al. 2026). Few tests of AI systems account for their adaption to dynamic data, which obfuscates such adaptive strategies. To detect and combat these risks, the DoW requires a universal approach to elicit, characterize, and compare the behavior of AI agents amid changing contexts (Li et al. 2026). Such a testbed shall rely only on queries and outputs of the subject AI system, rather than direct access to the model itself. Economic frameworks permit the measurement and comparison of decisions. As the basis of extensive prior research, models of auctions, markets, and other economic arenas provide a critical baseline of organic patterns of human behavior (e.g. Hausch 1986, Martinez-Saito 2019). Emerging open-source tools, such as Magentic Marketplace (Bansal et al. 2026) have shown promise in revealing behavioral variations in market-focused contexts. Furthermore, unlike existing assessments of AI risk, economic frameworks do not rely on a priori definitions of “harmful” or “helpful” traits (Vijayvargiya et al. 2026), and these testbeds allow for diverse social strategies such as deception and collaboration.
- Category
- R&D
- Industry
- AI & Data
- Technology
- AI / ML
- Target stage
- Needs verification
- Project duration
- Needs verification
- Estimated preparation
- Needs verification
Funding
Award size and how it is paid
- Award range
- Needs verification
- Currency
- USD
- Total programme budget
- Needs verification
- Support type
- Needs verification
- Co-funding
- Needs verification
- Matching fund
- Needs verification
- Disbursement
- Needs verification
- Note
- -
Eligibility
Can we actually apply?
SBIR/STTR 은 미국 중소기업만 지원할 수 있습니다(Small Business Act 법정 요건). • 계열사를 포함해 상시 종업원 500명 이하 • 미국 시민 또는 영주권자 1인 이상이 50%를 초과해 직접 소유·지배 • 미국 내 사업장을 두고 주로 미국 내에서 사업을 영위할 것 • 수행책임자(PI)의 주된 근무처가 신청 기업일 것 출처: https://www.sbir.gov/faq/eligibility-requirements
- Company age
- Needs verification
- Employees
- 0명 ~ 500명
- Revenue limits
- Needs verification
- Consortium
- Needs verification
Location Requirements
Geography and legal-entity conditions
- Primary country
- 🇺🇸 United States
- Also eligible
- None
- Foreign companies
- No
- Local entity
- Required
- Location condition
- 미국 내 사업장을 두고 미국 시민·영주권자가 50%를 초과해 소유한 중소기업만 신청할 수 있습니다(법정 요건).
Required Documents
Required/optional · issuer · difficulty · validity · cautions
Document requirements were not captured (needs verification). Check the official announcement.
Application Process
How you apply
Application process not captured.
Evaluation
Review process and criteria
Evaluation process not captured.
Timeline
Announcement → intake → review → agreement → execution
Process steps were not captured.
Equity & Financial Terms
Equity, loans and contract conditions
- Equity required
- Needs verification
- Loan
- Non-dilutive grant
- Duplicate funding
- Needs verification
- IP ownership
- Needs verification
- Deliverable ownership
- Needs verification
- Exclusivity
- Needs verification
- Right of first refusal
- Needs verification
- Audit & settlement
- Needs verification
Risk Intelligence
Should we apply at all? — five dimensions plus an overall score (lower is safer)
No risk assessment yet. Re-analyse from Admin.
IP / Idea Protection
How much of your technology and idea you must disclose
No IP assessment yet.
AI Analysis
Pros · cons · difficulty · competition · attractiveness
From paperwork and review stages
Heuristic from award size and type
Amount, terms and IP combined
Weighted average of five dimensions
- Needs verification
- Needs verification
- Nothing flagged
Similar Programs
Comparable by type, country and technology
- Ground and Air Launched Drone Swarms Create a Self-Protecting Perimeter Using Autonomous AI
🇺🇸 United States · R&D program · U.S. Department of War — USAF
Needs verificationD-14 - Agentic-AI, Schema-Driven Decision Management for Auditable Studies and Acquisition Decisions
🇺🇸 United States · R&D program · U.S. Department of War — ARMY
Needs verificationD-14 - Agentic AI Based Cognitive Radar for GEOINT Mission
🇺🇸 United States · R&D program · U.S. Department of War — OSD
Needs verificationD-14 - AI/ML for Next Generation of Missile Detection, Warning, Tracking, and Reporting
🇺🇸 United States · R&D program · U.S. Department of War — USAF
Needs verificationD-14 - Sustainable AI Technologies for Low Resource Environments
🇺🇸 United States · R&D program · National Science Foundation
Needs verificationD-273
Source & Provenance
Every value carries its source URL and fetch time
- Official URL
- https://www.dodsbirsttr.mil/topics-app/
- Application URL
- https://www.dodsbirsttr.mil/topics-app/
- Last fetched
- 2026.10.07
- Human review
- Not reviewed — check the official announcement
- First seen
- 2026.10.07
- Data status
- Auto-collected, not reviewed
- AI enrichment
- Needs verification
- PRIMARYSBIR.gov topic — Influence Benchmarks for AI Systems
Fetched 2h ago
- HTMLSBIR/STTR 자격요건 (법정)
Fetched 2h ago
- HTML발주 기관 공식 공고
Fetched 2h ago
- · 2h ago — 최초 수집 (9af45239)