SKILL: Bug Identification
Metadata
Description
Systematic bug identification methodology: source code review patterns, black-box testing strategies, taint analysis, dangerous function hunting, data flow tracing, and automated scanning setup. Use for code audits, bug bounty triage, or building vulnerability identification pipelines.
Trigger Phrases
Use this skill when the conversation involves any of:
bug identification, code review, taint analysis, dangerous functions, data flow, source audit, black box, vulnerability identification, static analysis, code audit, bug hunting
Instructions for Claude
When this skill is active:
- Load and apply the full methodology below as your operational checklist
- Follow steps in order unless the user specifies otherwise
- For each technique, consider applicability to the current target/context
- Track which checklist items have been completed
- Suggest next steps based on findings
Full Methodology
Bug Identification
Overview
Bug identification is the process of discovering potential vulnerabilities in software through various techniques including static analysis, dynamic analysis, and fuzzing. This document outlines methodologies and tools for effective vulnerability research.
For practical exploit development, see Exploit Development.
flowchart TD
BugId["Bug Identification"]
%% Main Methods
Static["Static Analysis"]
Dynamic["Dynamic Analysis"]
Fuzzing["Fuzzing"]
AI["AI-Assisted"]
%% Static Analysis Methods
CodeReview["Manual Code Review"]
RevEng["Reverse Engineering"]
PatchDiff["Patch Diffing"]
StaticTools["Static Analysis Tools"]
SBOM["Supply Chain Analysis"]
%% Dynamic Analysis Methods
DebugTrace["Debugging/Tracing"]
DBI["Dynamic Binary Instrumentation"]
Taint["Taint Analysis"]
SymExec["Symbolic Execution"]
Snapshot["Snapshot Analysis"]
%% Fuzzing Methods
DumbFuzz["Dumb Fuzzing"]
SmartFuzz["Smart Fuzzing"]
EvoFuzz["Evolutionary Fuzzing"]
LLMFuzz["LLM-Guided Fuzzing"]
%% AI Methods
LLMTriage["LLM Crash Triage"]
MLPattern["ML Pattern Recognition"]
AutoVariant["Automated Variant Analysis"]
%% Connections
BugId --> Static
BugId --> Dynamic
BugId --> Fuzzing
BugId --> AI
Static --> CodeReview
Static --> RevEng
Static --> PatchDiff
Static --> StaticTools
Static --> SBOM
Dynamic --> DebugTrace
Dynamic --> DBI
Dynamic --> Taint
Dynamic --> SymExec
Dynamic --> Snapshot
Fuzzing --> DumbFuzz
Fuzzing --> SmartFuzz
Fuzzing --> EvoFuzz
Fuzzing --> LLMFuzz
AI --> LLMTriage
AI --> MLPattern
AI --> AutoVariant
%% Combinations
Taint -.-> Fuzzing
SymExec -.-> Fuzzing
RevEng -.-> Fuzzing
AI -.-> Fuzzing
AI -.-> Static
class BugId primary
Vulnerability Research Methodology
Phase 1: Reconnaissance
- Target Enumeration: Identify version, dependencies, configuration
- Attack Surface Mapping: List all input vectors, APIs, protocols
- Documentation Review: RFCs, specifications, developer docs
- Prior Art Analysis: CVE database, exploit-db, bug trackers
Phase 2: Static Analysis
- Source Review: If available, focus on parsing/validation code
- Binary Analysis: Reverse engineering with Ghidra/IDA
- Patch Diffing: Compare vulnerable vs patched versions
- SBOM Analysis: Check third-party component vulnerabilities
Phase 3: Dynamic Analysis
- Behavioral Analysis: Monitor syscalls, network, file I/O
- Debugging: Trace execution paths with controlled input
- Instrumentation: Coverage-guided exploration
- Taint Analysis: Track input propagation
Phase 4: Fuzzing
- Corpus Generation: Create valid seed inputs
- Harness Development: Isolate target functionality
- Coverage Monitoring: Identify untested code paths
- Crash Triage: Classify and prioritize findings
Phase 5: Exploitation
- Primitive Development: Convert bug to reliable primitives
- Mitigation Bypass: Defeat ASLR, DEP, CFG, etc.
- Payload Development: Create working exploit
- Weaponization: Package for real-world use (if authorized)
Attack Surface Identification
Before diving into specific bug hunting techniques, it's essential to understand where to look for vulnerabilities.
Windows User Mode
- Shared Memory
- RPC
- Named Pipes
- File & Network IO
- Windows Messages
- For authentication-related vulnerabilities, see Windows Auth
Kernel
- Device Drivers
- Many third-party software with drivers to target
- Can accept arbitrary user input via the
IOCTL interface
- Also performs actions when we
open,close handles to it
- OS
- Drivers that handle hardware and user input
- Intercepts/transitions from user to kernel
- Modern Linux interfaces (hotspots)
- io_uring: SQE size/offset confusions, submission/completion race windows, kernel copy‑sizes derived from user buffers
- userfaultfd: cross‑thread write‑what‑where and TOCTOU primitives during fault handling
- seccomp user‑notifier: confused‑deputy patterns in broker processes; notifier time‑of‑check vs time‑of‑use gaps
- Hyper-V & VTL Interfaces – On many modern Windows 11 systems (especially 24H2 on supported hardware), Virtualization‑Based Security and VTL1 are enabled or easily enabled by policy. Treat the hypervisor surface (e.g.,
hvix64.exe and synthetic MSRs) as a common kernel target, and verify VBS/HVCI status on the host before assuming defaults.
Drivers
- DriverEntry: registers for any callbacks, setup structure, etc
- I/O Handlers: handlers that get called when a process attempts to
open,close,etc the driver, IOCTL allows driver functionality to be called from user processes
- Practical triage example (CVE‑2025‑8061):
- IOCTL handlers that accept a fixed‑size struct and pass a user‑controlled
PHYSICAL_ADDRESS directly to MmMapIoSpace
- then memcpy out/in mapped memory (sometimes via wrappers that swap src/dst) indicate physical memory read/write primitives.
- Similarly, unguarded MSR read/write paths yield
RDMSR/WRMSR primitives.
- See the Lenovo
LnvMSRIO.sys case study in windows-kernel.md
eBPF & XDP
- BPF helpers and verifier: pointer leaks, verifier bypass, JIT bugs
- User‑entry vectors:
bpf() syscall, privileged pods in Kubernetes, Cilium datapath
- Tooling:
bpftool, verifier logs, bpftrace scripts for quick triage
- CO‑RE skeletons (
bpftool gen skeleton) simplify packaging portable tracing probes.
- BPF LSM hooks allow low‑overhead coverage feedback on security‑critical kernel paths; export events with
trace_pipe.
Container & Micro‑VM Surface
- Namespace/cgroup escapes, device‑mapper abuse, races in snapshotting backends (e.g., overlayfs)
- Micro‑VM hypercalls in Firecracker, CloudHypervisor, Kata Containers
- For detailed container exploitation techniques, see Container
Cloud‑Native & IAM Bugs
- Misconfigured IAM policies, privilege‑escalating API actions (AWS
sts:AssumeRole, Azure Golden SAML)
- SSRF paths into metadata services (
169.254.169.254, IMDSv2 bypass techniques)
- Race conditions in managed control‑plane components (Kubernetes API server, AWS Lambda workers)
- Kubernetes Attack Vectors: look at kubernetes for a deeper checklist
- Serverless Vulnerabilities:
- Lambda layer poisoning
- Function URL authentication bypass
- Event injection through SQS/SNS/EventBridge
- Cold start race conditions
Network / Transport Protocol Parsers
- QUIC / HTTP/3: coalesced frames, reorder/timing corner cases; verify against RFC 9000 (QUIC) and RFC 9114 (HTTP/3)
- HTTP/2: stream state machine desync; flow‑control integer edge cases (RFC 7540)
- gRPC / Protobuf: length truncation across language FFI, map/list coercion; see gRPC framing and protobuf varint rules
- GraphQL: input coercion and resolver recursion limits; check GraphQL spec for type coercion semantics
WebAssembly Runtimes
- WASM JIT optimization bugs in V8, Wasmtime, Wasmer
- WASI sandbox escapes through host‑call interfaces
- Typed‑Func‑Refs, GC, Tail‑calls, Memory64 expand type/bounds confusion surface. See the WebAssembly proposals status page for current rollout and engine adoption.
- Checklist:
- validate table element types/import signatures/hostcall marshalling
- fuzz mixed 32/64-bit memories.
- Fuzzing tip: compile native libs to WASM for fast, deterministic mutation cycles
Browser / JS Engine Exploitation
Modern V8 Architecture (2024-2025)
V8 now uses a multi-tier JIT pipeline with distinct exploitation characteristics:
- Ignition (Interpreter): Bytecode interpreter; rarely targeted directly
- Maglev (Mid-tier JIT): Introduced Chrome 115+; simpler IR than TurboFan
- TurboFan (Optimizing JIT): Aggressive optimization; traditional exploitation target
- Turboshaft: New IR replacing TurboFan internals; different optimization patterns create new bug classes
- Type lattice changes affecting confusion bugs
- Maglev → Turboshaft transition paths expose state inconsistencies
- Node-based to block-based IR transition
V8 Maglev Exploitation
- Integer overflow in Maglev's fast-path arithmetic
- Corrupted HeapNumber backing store via Maglev bounds check bypass
- Map/ElementsKind confusion in polymorphic inline caches
WebAssembly JSPI (JavaScript Promise Integration)
- Stack Heap Spray: Suspended WASM stacks allocated on heap; predictable layout
- Type Confusion:
WebAssembly.Suspending wrapper type mismatch
- Info Leak: Stack pointers exposed through Promise resolution chains
- Sandbox Escape: JSPI bridges JS/WASM boundary; bypass traditional WASM isolation
Spectre-BHB Browser Mitigations
- Chrome 120+: Site Isolation per-frame; shared array buffer restrictions
- Firefox 122+: Process-per-site with BHI fences in JIT trampolines
- Safari 17.4+: WebKit JIT speculation guards on type checks
Site Isolation Plus
- Frame-level process isolation: Each cross-origin frame in separate process
- Cross-origin memory protection: Hardware-backed memory isolation
- New IPC attack surface: Mojo interface exploitation required for escapes
- Renderer → Browser requirements: Need Mojo race or type confusion
- New Info-Leak Requirements:
- Traditional
SharedArrayBuffer + Atomics timing attacks less reliable
- Need alternative side-channels: CSS timing, WebGL shader execution, AudioContext
- Cross-origin info leaks require chaining multiple primitives
Practical Browser Exploitation Workflow
Target Selection:
- V8 Maglev for Chrome/Edge (faster development cycle = more bugs)
- JSC for Safari (less scrutiny than V8)
- SpiderMonkey for Firefox (IonMonkey/Warp still viable)
Primitive Development:
addrof: Leak object addresses (info leak)
fakeobj: Craft fake object (type confusion)
arbread/arbwrite: Arbitrary memory access
shellcode: RWX page or WASM JIT abuse
Sandbox Escape:
- Mojo IPC race conditions
- GPU process exploitation via WebGL
- Utility process TOCTOU (Chrome's new architecture)
Post-Exploitation:
- Chrome: Target browser process via Mojo
- Safari: XPC service exploitation for sandbox escape
- Firefox: Target parent process via IPC
Firmware & Embedded
- UEFI DXE driver flaws, BMC web console auth bypass, ECU/CAN message injection
- BLE & Zigbee stack overflows, heap exploits in
btstack, lwIP
macOS / Apple‑Silicon Kernel
- IOKit user‑client input validation, IOMFB allocator corner‑cases
- Hypervisor.framework fuzzing with
hv_fuzz
Mobile Platforms (iOS/Android)
iOS 17+ Exploitation
- PAC Bypass: Pointer Authentication Code bypass via signing gadgets
- PPL Bypass: Page Protection Layer exploitation for kernel r/w
- Secure Enclave: SEP exploitation via malformed Mach messages
- Neural Engine: ANE kernel driver attack surface
Android 14+ Exploitation
- MTE (Memory Tagging): Probabilistic bypass with tag collisions
- GKI (Generic Kernel Image): Vendor hooks as attack surface
- Scudo Hardening: Heap exploitation with hardened allocator
- Hardware Attestation: Keymaster/StrongBox TEE attacks
Cross-Platform Mobile
- Flutter: Dart VM type confusion, FFI boundary issues
- React Native: JavaScript bridge serialization bugs
- Unity: IL2CPP memory corruption, native plugin vulnerabilities
Supply Chain Attack Surface
Package Manager Vulnerabilities
- Dependency Confusion: Internal vs public package name conflicts
- Typosquatting: Similar package names (numpy vs numpi)
- Manifest Manipulation: Lock file poisoning, version pinning bypass
- Build-time Injection: Malicious install scripts, post-install hooks
CI/CD Pipeline Analysis
- GitHub Actions: Workflow poisoning via PR from forked repos
- Jenkins: Groovy script injection, plugin vulnerabilities
- Docker: Build argument exploitation, base image substitution
- Secrets Exposure: Environment variables in build logs, artifact leakage
AI & LLM Application Security
- Prompt‑injection, sandbox boundary escapes, hidden‑channel data exfil
- See AI Security for a deeper checklist
Confidential‑Computing / TEE Surface
- Intel TDX: diff
tdx.ko or tdx_psci.c between kernel LTS branches to spot new GPA→HPA validation checks.
- AMD SEV‑SNP: look for unchecked
VMGEXIT leafs in PSP firmware; sevtool --decode helps locate IDA entry points.
- Arm CCA / RMM: analyze SMC handlers inside Realm Management Monitor (RMM) EL3 firmware.
- Cloud offerings (Azure CCE, Google C3): focus on paravirtualised MMIO and attestation report flows exposed to guests.
- For TEE-specific exploitation, see Secure Enclaves
GPU & vGPU Surface
- HGX HMC (verify CVE/advisories): research indicates malformed NVLINK‑C2C packets can corrupt HMC register space; confirm against vendor advisories for the specific platform.
- vGPU manager IOCTLs: diff
nvidia‑vgpu‑mgr monthly; watch VGPU_PLUGIN_IOCTL_GET_STATE and similar calls for unchecked buffers.
- LeftoverLocals info‑leak: contiguous VRAM allocations can leak data from prior tenants in multi‑tenant AI clusters.
Hardware Security Attack Surface
Side-Channel Analysis
- Power Analysis: DPA/SPA attacks on cryptographic operations
- Electromagnetic (EM): Near-field probing of processor emissions
- Timing Attacks: Cache timing, branch prediction analysis
- Acoustic: Key extraction via CPU sound emissions
Fault Injection
- Voltage Glitching: Brown-out attacks on secure boot
- Clock Glitching: Skip instruction execution
- Laser Fault Injection (LFI): Targeted bit flips
- EM Pulse Injection: Wider area fault induction
Hardware Implants & Supply Chain
- PCB Modification: Added components, trace rerouting
- Firmware Backdoors: UEFI/BMC persistent implants
- Hardware Trojans: Malicious logic in ICs
- DMA Attacks: PCIe, Thunderbolt, FireWire exploitation
EDR Driver Vulnerability Research
Common vulnerability types in EDR drivers
- Authorization bypass issues
- Memory corruption in IOCTL handlers
- Race conditions in driver communication
- Improper input validation
- For detailed EDR analysis techniques, see EDR
Research methodology
- Identify accessible driver interfaces
- Reverse engineer IOCTL/message handlers
- Analyze authorization mechanisms
- Test for input validation flaws
- Look for race conditions and memory corruption
Tools for driver analysis
- IDA Pro / Ghidra for reverse engineering
- WinDbg for dynamic analysis
- Process Monitor for behavior analysis
- Custom fuzzing tools for interface testing
Quick triage rubric (post‑crash)
- Buffer overflow vs UAF: check access type, allocation lifetime, and red‑zones (ASan/KASAN reports)
- Integer issues: trace size/length and allocation math; look for truncation/casts
- Logic bugs: unexpected state transitions without memory errors; validate auth/flags
- Info‑leaks: uninitialized reads, OOB reads, pointer/string formatters
Coverage‑first recon checklist
- Produce one baseline coverage run (e.g.,
drcov, Intel® PT, or Lighthouse import)
- Identify cold paths reachable from attacker inputs
- Seed corpus: include minimal valid examples that traverse target parsers
- Enable lightweight oracles (ASan/UBSan/KASAN) where feasible to maximize signal
Static Analysis Methods
Static analysis examines code without execution to identify potential vulnerabilities.
Manual Code Review
- Installing the target application and examining its structure
- Enumerating the ways to feed input to it
- Examine the file formats and network protocols that the application uses
- Locating logical vulnerabilities or memory corruptions
- For Windows-specific techniques, see Windows Kernel
- For Linux-specific techniques, see Linux
Patch Diffing
Patch diffing compares vulnerable and patched versions of binaries to identify security changes.
What is Patch Diffing
Patch diffing is a technique to identify changes across versions of binaries related to security patches. It compares a vulnerable version of a binary with a patched one to highlight the changes, helping to discover new, missing, and interesting functionality across versions.
Benefits
- Single Source of Truth: Without a CVE blog post or sample POC, a patch diff can be the only source of information to determine changes and deduce the original issue.
- Vulnerability Discovery: While understanding the original issue, you may discover additional vulnerabilities in the troubled code area.
- Skill Development: Patch diffing provides focused practice in reverse engineering and helps build mental models for various vulnerability classes.
Challenges
- Asymmetry: Small source code changes can drastically affect compiled binaries.
- Finding Security-Related Changes: Security patches often include other changes like new features, bug fixes, and performance improvements.
- Minimizing Noise:
- Diff the correct binaries to avoid analyzing unrelated updates
- Reduce the time delta between compared versions
- Use binary symbols when available to add precision to comparisons
Tools
- IDA Pro with plugins like DarunGrim and Diaphora
- BinDiff Works with analysis output from IDA or Ghidra
- Ghidriff: Ghidra binary diffing engine
- Radare2 (radiff2)
- Ghidra Version Tracking Tool
- Ghidra 11 built-in Partial Match Correlator
Patch Diffing Workflow
The process of patch diffing typically follows these steps:
Preparation
- Create a diffing session
- Load binary versions (vulnerable and patched)
- Ensure binaries pass preconditions
- Run auto-analysis on both binaries
Evaluation
- Run correlators to find similarities
- Generate associations between binaries
- Evaluate matches between functions
- Accept matching functions
- Analyze differences until sufficient understanding is reached
Function Analysis
- Identify new functions: Functions in the patched binary with no match in the original
- Identify deleted functions: Functions in the original binary with no match in the patched version
- Identify changed functions: Functions that exist in both versions but have been modified
- Focus on functions with security relevance (often indicated by their names or based on CVE descriptions)
Interpreting Results
- New functions often indicate added security checks or validation
- Changed functions may show modified logic for handling edge cases
- Correlate changes with public CVE information when available
- Remember that patches are not necessarily atomic - multiple issues may be fixed in one update
When using Ghidra's Version Tracking:
- Use "Show Only Unmatched Functions" filter to identify new or deleted functions
- Look for functions with a similarity score below 1.0 to find modified functions
- Examine the modified functions to understand what security checks were added
Starting with Ghidra 11 (December 2024) a built-in Partial Match Correlator covers most PatchDiffCorrelator use-cases; install the plugin only if you need bulk-mnemonics scoring.
Case Study: 7‑Zip Symlink Path Traversal
- Target: 7‑Zip 24.09 (vulnerable) → 25.00 (fixed)
- File of interest:
CPP/7zip/UI/Common/ArchiveExtractCallback.cpp
- High‑signal edits: absolute‑path detection and link‑path validation for WSL/Linux symlinks converted on Windows.
Minimal security‑relevant diff (simplified):
-bool IsSafePath(const UString &path)
+static bool IsSafePath(const UString &path, bool isWSL)
{
CLinkLevelsInfo levelsInfo;
- levelsInfo.Parse(path);
+ levelsInfo.Parse(path, isWSL);
return !levelsInfo.IsAbsolute
&& levelsInfo.LowLevel >= 0
&& levelsInfo.FinalLevel > 0;
}
+bool IsSafePath(const UString &path);
+bool IsSafePath(const UString &path)
+{
+ return IsSafePath(path, false); // isWSL
+}
-void CLinkLevelsInfo::Parse(const UString &path)
+void CLinkLevelsInfo::Parse(const UString &path, bool isWSL)
{
- IsAbsolute = NName::IsAbsolutePath(path);
+ IsAbsolute = isWSL ? IS_PATH_SEPAR(path[0]) : NName::IsAbsolutePath(path);
LowLevel = 0;
FinalLevel = 0;
}
Root cause (logic):
- Linux/WSL symlink data containing a Windows‑style path (e.g.,
C:\...) was treated as relative by the Linux absolute‑path check, setting linkInfo.isRelative = true.
SetFromLinkPath prefixed the symlink’s zip‑internal directory when building relatPath, letting IsSafePath(relatPath) pass despite an absolute Windows target.
- A subsequent “dangerous link” guard checked
_item.IsDir; non‑directory symlinks skipped the validation.
- Result: symlink creation to arbitrary absolute Windows paths; extracted files written into the link target.
Practical triage checklist:
- Search this file for:
IsSafePath, CLinkLevelsInfo::Parse, SetFromLinkPath, CloseReparseAndFile, FillLinkData, CLinkInfo::Parse, _ntOptions.SymLinks_AllowDangerous.
- Verify absolute‑path detection across OS semantics (Linux vs Windows) and that relative/absolute status cannot be desynced by mixed‑style paths.
- Ensure “dangerous link” checks run for both files and directories; avoid
_item.IsDir short‑circuiting validation for file symlinks.
- Confirm
IsSafePath evaluates the final target path after concatenations; normalize before validation.
Quick repro (Windows, developer mode or elevated):
- Create zip structure:
data/link → symlink to C:\Users\<USER>\Desktop
data/link\calc.exe → payload file
- If
link is extracted first, subsequent writes follow the symlink into the absolute target directory.
Apple Patch Diffing
- Identify a CVE of interest
- Download corresponding IPSW and update and N-1
- use ipsw.me to download those files
- convert the downloaded
.ipsw to .zip
- Determine changes for update
- Map binaries to CVE
- Extract the related file(s)
- Diff the binaries
- Root cause the vulnerability
# Downloading the correct IPSWs
ipsw download --device Macmini9,1 -V -b 23A344
ipsw download --device Macmini9,1 -V -b 23B74
# Comparing two different IPSWs
ipsw diff UniversalMac_14.0_23A344.ipsw UniversalMac_14.1_23B74.ipsw
# What is inside the DSC
ipsw extract -d IPSW
# Extracting files
ipsw extract -f -p file
# Extracting specific architecture file
ipsw macho lipo Contacts
Windows 11 Patch Diffing
This checklist mirrors the Apple IPSW workflow but uses Microsoft tooling and build numbers.
Identify the target update
- Open Settings → Windows Update → Update history or consult the Windows Release Health dashboard to note the KB and OS build numbers (e.g., KB5037778 → build 22631.3525).
- Record the previous build you want to diff against (e.g., 22631.3447).
Collect the binaries
winbindex download tcpip.sys 10.0.22631.3447 10.0.22631.3525
mkdir pre,post
wget -Uri https://www.catalog.update.microsoft.com/Download.aspx?q=KB5037778 -OutFile kb.msu
expand -F:* .\kb.msu .\post
# repeat for the older KB into .\pre
# Download both *UUP* bundles, then run
uup_download_windows.cmd --extract
# and copy changed PE files to *pre* / *post*
- Fetch matching symbols
# Requires Debugging Tools for Windows
foreach ($ver in '3447','3525') {
symchk /r .\$ver /s SRV*https://msdl.microsoft.com/download/symbols
}
Load in the disassembler
- Open tcpip.sys from both pre and post folders in IDA 8+ or Ghidra 11; ensure PDB symbols resolve.
- Save the IDA databases (e.g.,
tcpip_3447.i64, tcpip_3525.i64).
Run the diff
Triage the results
- Sort by Similarity % ascending; investigate anything below 95 %.
- Focus on functions with names like
Validate, Parse, Copy, Check, or protocol‑specific handlers (IppReceiveEsp, Ipv6pFragmentReassemble, etc.).
- Determine whether changes add bounds checks, size validations, or privilege checks.
Validate in a lab VM
- Snapshot two Windows 11 VMs (build 3447 and 3525).
- Attach WinDbg (kernel mode) using
bcdedit /dbgsettings net hostip:<IP> port:<PORT>.
- Reproduce the issue against the pre‑patch VM; confirm no crash or breakpoint triggers in the post‑patch VM.
Automate monthly
- Schedule a PowerShell script that, every Patch Tuesday (second Tuesday), downloads the latest Cumulative Update, extracts changed PE files, retrieves symbols, and launches a headless Diaphora diff.
- Email the generated HTML report to quickly spot new attack surface.
[!TIP]
For large modules like ntoskrnl.exe, diff only the .text section to save RAM:
bindiff --primary ntoskrnl_pre.i64 --secondary ntoskrnl_post.i64 --section .text
Linux Kernel Patch Diffing
Patch‑diffing Linux kernels is often faster at the source level, but for binary‑only targets (vendor kernels, modules) function‑level diffing is still practical.
Identify target builds
- Note distro and kernel build (e.g., Ubuntu
6.8.0-47-generic, RHEL 5.14.0-503).
- Capture both pre and post versions (package changelogs or CVE bulletins help).
Fetch kernel images and debug info
- Ubuntu/Debian:
# Discover versions
apt list -a linux-image-generic | cat
# Download image + modules dirs (repeat for both versions)
apt-get download linux-image-unsigned-<ver>-generic linux-modules-<ver>-generic
# Debug symbols via debuginfod (preferred to ddebs)
export DEBUGINFOD_URLS="https://debuginfod.ubuntu.com https://debuginfod.debian.net"
- Fedora/RHEL/CentOS:
dnf download kernel-core-<ver> kernel-debuginfo-<ver>
rpm2cpio kernel-core-<ver>.rpm | cpio -idmv
rpm2cpio kernel-debuginfo-<ver>.rpm | cpio -idmv
Extract vmlinux
# If only vmlinuz is present, use the upstream helper
/usr/src/linux-headers-<ver>/scripts/extract-vmlinux /boot/vmlinuz-<ver> > vmlinux-<ver>
# Or take vmlinux directly from debuginfo package tree
Identify changed modules quickly
# Compare module trees (pre vs post)
rsync -rcn --delete /lib/modules/<pre>/ /lib/modules/<post>/ | grep -E "\.ko$" | sed 's/^/chg: /'
Function‑level binary diff
- Open
vmlinux-<pre> and vmlinux-<post> in Ghidra 11/IDA 8 and run Diaphora/BinDiff/Ghidriff.
- For hot subsystems (e.g.,
io_uring, net/ipv6, fs/overlayfs), diff only the relevant .ko pairs to reduce noise.
Source‑level triage (when sources are available)
# Ubuntu example: unpack both source trees, then
git diff --no-index -- function.c.orig function.c.patched | less
# Or use diffoscope for enriched reports
Symbolization and crash mapping (cheat‑sheet)
# Decode kernel oops backtraces to lines
./scripts/decode_stacktrace.sh vmlinux /lib/modules/<ver>/build < dmesg.log
# Map PC to file:line quickly
addr2line -e vmlinux-<ver> 0xffffffff81234567
[!TIP]
For modern distros built with Clang: KCFI and fine‑grained CFI thunks create many small stub changes; filter by real function body deltas to focus on security‑relevant logic.
[!NOTE]
Syzkaller routinely bisects kernel bugs; consult syzbot reports for reproducers and fix commits, then confirm your diff isolates the same region before deeper RE.
Kernel network parser identification heuristics (SMB2-inspired, broadly applicable)
Cross-field invariants (length/offset/next)
- Always validate
(offset + length) <= remaining_buffer and <= total_buffer using a widened type (e.g., u64) before arithmetic; reject on overflow with check_add_overflow()/array_size() helpers.
- For chained entries with a
next field, assert: next >= sizeof(entry_header), next <= remaining_buffer, and that pointer advancement actually makes progress. For entries carrying sub-lengths (e.g., name_len, value_len), assert header + name_len + value_len <= next.
- Do not cast to a struct until the full header is present and aligned; gate recasts with a prior
buf_len check.
Fixed-size buffers vs variable-length payloads
- Ban unbounded copies/crypto/decompression into fixed-size arrays. Require
len <= sizeof(array) (or clamp with min_t() and bail) when writing into in-struct arrays.
- Crypto transforms are just writes with extra steps: if using ARC4/AES helpers that copy
len bytes into a fixed buffer (e.g., session keys), bound len against a named maximum constant and prefer allocating a buffer sized from validated len.
Type/width hazards
- Normalize parser math to a wide unsigned type before comparisons; avoid truncating
u32/u64 fields into u16 for size checks. Favor size_t/u64 for offset+len arithmetic, then compare to buf_len of the same width.
Loop structure around next
- Pattern to flag:
e = (struct entry *)((char *)e + next); without a preceding block that revalidates buf_len and the entry’s internal sub-lengths.
- Ensure a break condition on exhaustion and reject zero/negative progress values to avoid infinite loops or pointer stagnation.
Allocation-size correlation
- When parser-controlled
len influences a subsequent write into an object from a fixed SLUB cache (e.g., kmalloc-512), ensure the write length is bounded by the destination object field, not just the incoming length.
Patch-diff signals to prioritize
- Newly added guards like
if (len > CONST) return -EINVAL;, if (buf_len < sizeof(struct foo)) return -EINVAL;, or conversions to min_t(size_t, len, sizeof(...)).
- Insertions of
check_add_overflow(offset, len, &sum) or array_size(n, sz) helpers in hot parse paths.
Static query seeds (Semgrep/CodeQL), to tune per codebase
- Unbounded copies into struct fields:
rules:
- id: c-fixed-array-unbounded-copy
languages: [c, cpp]
patterns:
- pattern: memcpy($DST, $SRC, $LEN)
- pattern-inside: |
struct $S { ... char $BUF[$N]; ... };
...
$DST = &...->$BUF
- pattern-not: memcpy($DST, $SRC, MIN($LEN, sizeof(*$DST)))
message: Unbounded copy into fixed-size struct field
severity: WARNING
- Dangerous
next-driven pointer arithmetic without bounds checks:- id: c-parser-next-missing-bounds
languages: [c, cpp]
pattern: |
$E = (struct $T *)((char *)$E + $NEXT);
message: Parser advances by user-controlled 'next' without prior buf_len/sizeof checks
severity: WARNING
- Crypto/decompression writes to fixed arrays (seed with function names in your tree, e.g.,
*_crypt, *_decrypt, decompress_*).
Dynamic confirmation (cheap)
- Grammar fuzz small invariants: send
next < header, next > remaining, name_len + value_len > next, and len > MAX_CONST variants; expect -EINVAL/reject. If not, investigate.
- Use TUN/TAP + KCOV to drive packet/SMB request paths; enable KASAN/KMSAN to surface overflows/leaks early.
Reference (motivating example)
- Lessons distilled from a 2025 ksmbd remote chain writeup combining a fixed-buffer overflow in NTLM auth with an EA
next validation issue — see Will’s Root: Eternal‑Tux: KSMBD 0‑Click RCE (https://www.willsroot.io/2025/09/ksmbd-0-click.html).
Case Study: EvilESP Vulnerability (CVE-2022-34718)
This case study demonstrates real-world patch diffing to identify a Windows TCP/IP RCE vulnerability.
Vulnerability Overview
- CVE-2022-34718: Critical RCE in
tcpip.sys discovered in September 2022
- An unauthenticated attacker could send specially crafted IPv6 packets to Windows nodes with IPsec enabled
- Affects the handling of ESP (Encapsulating Security Payload) packets in IPv6 fragmentation
Patch Diffing Process
Binary Acquisition
- Used Winbindex to obtain sequential versions of
tcpip.sys (pre-patch and post-patch)
- Loaded both files in Ghidra with PDB symbols
Diff Analysis
- Used BinDiff to compare the binaries
- Identified only two functions with less than 100% similarity:
IppReceiveEsp and Ipv6pReassembleDatagram
Code Analysis
- Ipv6pReassembleDatagram: Added bounds check comparing
nextheader_offset against the header buffer length
- IppReceiveEsp: Added validation for the Next Header field of ESP packets
Root Cause Identification
- Found an out-of-bounds 1-byte write vulnerability
- ESP Next Header field is located after the encrypted payload data
- A malicious packet could cause
nextheader_offset to exceed the allocated buffer size
(Update: Server 2022 build 20349.2300, May 2024, hardened this code path; the original PoC needs a 2-byte pad tweak to reproduce the crash.)
Exploitation
- Required setting up IPsec security association on the victim
- Created fragmented IPv6 packets encapsulated in ESP
- Controlled the offset of the out-of-bounds write through payload and padding size
- Value written is controllable via the Next Header field
- Limited to writing to addresses that are 4n-1 aligned (where n is an integer)
- Initially achieved DoS with potential for RCE through further exploitation
Lessons Learned
- Binary patch diffing effectively identified the vulnerability location and nature
- Understanding protocol specifications (ESP and IPv6 fragmentation) was critical
- Simple buffer checks are still overlooked in complex networking code
- Even limited primitives (single byte overwrite at constrained offsets) can be dangerous
- For modern exploitation techniques, see Modern Samples
- For mitigation bypass techniques, see Modern Mitigations
When applying patch diffing to networking protocols:
- Understand the protocol specifications thoroughly
- Look for missing bounds checks in data processing
- Pay attention to buffer size calculations
- Check for proper validation of protocol field values and locations
- Consider evasion techniques for exploit deployment - see EDR
- Specs: ESP (RFC 4303) and IPv6 (RFC 8200) are essential references when reasoning about header placement and bounds
Semi-Automatic Patch Diffing
Manual Patch Diffing
- Microsoft releases patches on the second Tuesday of each month
- For Windows you can go to update catalog and search for the product version (for example
2022-10 x64 "Windows 10" 22H2)
- Try to look for smaller updates
mkdir 2022-09
mv *.msu 2022-09
cd 2022-09
mkdir extract
mkdir patch
expand -F:* .\*.msu .\extract
expand -F:* .\extract\<largest>.cab .\patch
expand -F:* .\patch\<largest>.cab .\patch
expand -F:* .\patch\Cab_* .\patch\
You can use Patch Extract instead
gci -Recurse c:\windows\WinSxS\ -Filter ntdll.dll
# copy the biggest file somewhere
.\delta_patch.py -i .\NTDLL\ntdll.dll -o ntdll.2020-10.dll .\NTDLL\r\ntdll.dll .\2020-10\x64\ntdll_<stuff>\f\ntdll.dll
.\delta_patch.py -i .\NTDLL\ntdll.dll -o ntdll.2020-11.dll .\NTDLL\r\ntdll.dll .\2020-11\x64\ntdll_<stuff>\f\ntdll.dll
Open unpatched version in IDA as the primary and the second, after that use BinDiff add-on to find the differences between them
then right click on a different matched function and see the visual diff in bin diff
also you can uncheck proximity browsing to see the entire function
look at red blocks and then yellow blocks
With patch clean script you can only see the actual changed files
Static Analysis Tools
IDA Pro and Rust Tools for Vulnerability Research
rhabdomancer: IDA Pro headless plugin that locates calls to potentially insecure API functions in binary files
- Helps auditors backtrace from candidate points to find pathways allowing access from untrusted input
- Generates JSON/SARIF reports containing vulnerable function calls and their details
- Written in Rust using IDA Pro 9 idalib and Binarly's idalib Rust bindings
haruspex: IDA Pro headless plugin that extracts pseudo-code generated by IDA Pro's decompiler
- Exports pseudo-code in a format suitable for IDEs or static analysis tools like Semgrep/weggli
- Creates individual files for each function with their pseudo-code
- Can be used as a library by third-party crates
augur: IDA Pro headless plugin that extracts strings and related pseudo-code from binary files
- Stores pseudo-code of functions that reference strings in an organized directory tree
- Helps trace how strings are used within the application
- Complements other reverse engineering tools
Modern Static Analysis Tools
- Semgrep Pro – Cloud-augmented SAST with custom rule sharing, LLM-assisted rule writing
- CodeQL – GitHub's semantic code analysis, excellent for variant analysis
- Weggli – Fast semantic search for C/C++ (better than grep for code patterns)
- Joern – Code property graph analysis for vulnerability discovery
- Ghidra 11.2+ – Built-in ML-powered function signature recognition
- Binary Ninja 4.0 – Cloud collaboration, improved HLIL decompilation
…(truncated)
1---2name: skill-bug-identification3description: Skill Bug Identification4---5# SKILL: Bug Identification67## Metadata8- **Skill Name**: bug-identification9- **Folder**: offensive-bug-identification10- **Source**: https://github.com/SnailSploit/offensive-checklist/blob/main/bug-identification.md1112## Description13Systematic bug identification methodology: source code review patterns, black-box testing strategies, taint analysis, dangerous function hunting, data flow tracing, and automated scanning setup. Use for code audits, bug bounty triage, or building vulnerability identification pipelines.1415## Trigger Phrases16Use this skill when the conversation involves any of:17`bug identification, code review, taint analysis, dangerous functions, data flow, source audit, black box, vulnerability identification, static analysis, code audit, bug hunting`1819## Instructions for Claude2021When this skill is active:221. Load and apply the full methodology below as your operational checklist232. Follow steps in order unless the user specifies otherwise243. For each technique, consider applicability to the current target/context254. Track which checklist items have been completed265. Suggest next steps based on findings2728---2930## Full Methodology3132# Bug Identification3334## Overview3536Bug identification is the process of discovering potential vulnerabilities in software through various techniques including static analysis, dynamic analysis, and fuzzing. This document outlines methodologies and tools for effective vulnerability research.3738For practical exploit development, see [Exploit Development](/exploit/development.md).3940```mermaid41flowchart TD42 BugId["Bug Identification"]4344 %% Main Methods45 Static["Static Analysis"]46 Dynamic["Dynamic Analysis"]47 Fuzzing["Fuzzing"]48 AI["AI-Assisted"]4950 %% Static Analysis Methods51 CodeReview["Manual Code Review"]52 RevEng["Reverse Engineering"]53 PatchDiff["Patch Diffing"]54 StaticTools["Static Analysis Tools"]55 SBOM["Supply Chain Analysis"]5657 %% Dynamic Analysis Methods58 DebugTrace["Debugging/Tracing"]59 DBI["Dynamic Binary Instrumentation"]60 Taint["Taint Analysis"]61 SymExec["Symbolic Execution"]62 Snapshot["Snapshot Analysis"]6364 %% Fuzzing Methods65 DumbFuzz["Dumb Fuzzing"]66 SmartFuzz["Smart Fuzzing"]67 EvoFuzz["Evolutionary Fuzzing"]68 LLMFuzz["LLM-Guided Fuzzing"]6970 %% AI Methods71 LLMTriage["LLM Crash Triage"]72 MLPattern["ML Pattern Recognition"]73 AutoVariant["Automated Variant Analysis"]7475 %% Connections76 BugId --> Static77 BugId --> Dynamic78 BugId --> Fuzzing79 BugId --> AI8081 Static --> CodeReview82 Static --> RevEng83 Static --> PatchDiff84 Static --> StaticTools85 Static --> SBOM8687 Dynamic --> DebugTrace88 Dynamic --> DBI89 Dynamic --> Taint90 Dynamic --> SymExec91 Dynamic --> Snapshot9293 Fuzzing --> DumbFuzz94 Fuzzing --> SmartFuzz95 Fuzzing --> EvoFuzz96 Fuzzing --> LLMFuzz9798 AI --> LLMTriage99 AI --> MLPattern100 AI --> AutoVariant101102 %% Combinations103 Taint -.-> Fuzzing104 SymExec -.-> Fuzzing105 RevEng -.-> Fuzzing106 AI -.-> Fuzzing107 AI -.-> Static108109 class BugId primary110```111112## Vulnerability Research Methodology113114### Phase 1: Reconnaissance115116- **Target Enumeration:** Identify version, dependencies, configuration117- **Attack Surface Mapping:** List all input vectors, APIs, protocols118- **Documentation Review:** RFCs, specifications, developer docs119- **Prior Art Analysis:** CVE database, exploit-db, bug trackers120121### Phase 2: Static Analysis122123- **Source Review:** If available, focus on parsing/validation code124- **Binary Analysis:** Reverse engineering with Ghidra/IDA125- **Patch Diffing:** Compare vulnerable vs patched versions126- **SBOM Analysis:** Check third-party component vulnerabilities127128### Phase 3: Dynamic Analysis129130- **Behavioral Analysis:** Monitor syscalls, network, file I/O131- **Debugging:** Trace execution paths with controlled input132- **Instrumentation:** Coverage-guided exploration133- **Taint Analysis:** Track input propagation134135### Phase 4: Fuzzing136137- **Corpus Generation:** Create valid seed inputs138- **Harness Development:** Isolate target functionality139- **Coverage Monitoring:** Identify untested code paths140- **Crash Triage:** Classify and prioritize findings141142### Phase 5: Exploitation143144- **Primitive Development:** Convert bug to reliable primitives145- **Mitigation Bypass:** Defeat ASLR, DEP, CFG, etc.146- **Payload Development:** Create working exploit147- **Weaponization:** Package for real-world use (if authorized)148149## Attack Surface Identification150151Before diving into specific bug hunting techniques, it's essential to understand where to look for vulnerabilities.152153### Windows User Mode154155- Shared Memory156- RPC157- Named Pipes158- File & Network IO159- Windows Messages160- For authentication-related vulnerabilities, see [Windows Auth](/exploit/windows-auth.md)161162### Kernel163164- _Device Drivers_165 - Many third-party software with drivers to target166 - Can accept arbitrary user input via the `IOCTL` interface167 - Also performs actions when we `open,close` handles to it168- _OS_169 - Drivers that handle hardware and user input170 - Intercepts/transitions from user to kernel171- _Modern Linux interfaces (hotspots)_172 - **io_uring**: SQE size/offset confusions, submission/completion race windows, kernel copy‑sizes derived from user buffers173 - **userfaultfd**: cross‑thread write‑what‑where and TOCTOU primitives during fault handling174 - **seccomp user‑notifier**: confused‑deputy patterns in broker processes; notifier time‑of‑check vs time‑of‑use gaps175- _Hyper-V & VTL Interfaces_ – On many modern Windows 11 systems (especially 24H2 on supported hardware), Virtualization‑Based Security and VTL1 are enabled or easily enabled by policy. Treat the hypervisor surface (e.g., `hvix64.exe` and synthetic MSRs) as a common kernel target, and verify VBS/HVCI status on the host before assuming defaults.176177### Drivers178179- _DriverEntry_: registers for any callbacks, setup structure, etc180- _I/O Handlers_: handlers that get called when a process attempts to `open,close,etc` the driver, `IOCTL` allows driver functionality to be called from user processes181- Practical triage example (CVE‑2025‑8061):182 - IOCTL handlers that accept a fixed‑size struct and pass a user‑controlled `PHYSICAL_ADDRESS` directly to `MmMapIoSpace`183 - then memcpy out/in mapped memory (sometimes via wrappers that swap src/dst) indicate physical memory read/write primitives.184 - Similarly, unguarded MSR read/write paths yield `RDMSR/WRMSR` primitives.185- See the Lenovo `LnvMSRIO.sys` case study in [windows-kernel.md](/exploit/windows-kernel.md)186187### eBPF & XDP188189- **BPF helpers and verifier**: pointer leaks, verifier bypass, JIT bugs190- **User‑entry vectors**: `bpf()` syscall, privileged pods in Kubernetes, Cilium datapath191- **Tooling**: `bpftool`, verifier logs, `bpftrace` scripts for quick triage192- **CO‑RE skeletons** (`bpftool gen skeleton`) simplify packaging portable tracing probes.193- **BPF LSM** hooks allow low‑overhead coverage feedback on security‑critical kernel paths; export events with `trace_pipe`.194195### Container & Micro‑VM Surface196197- Namespace/cgroup escapes, device‑mapper abuse, races in snapshotting backends (e.g., overlayfs)198- Micro‑VM hypercalls in Firecracker, CloudHypervisor, Kata Containers199- For detailed container exploitation techniques, see [Container](/exploit/container.md)200201### Cloud‑Native & IAM Bugs202203- Misconfigured IAM policies, privilege‑escalating API actions (AWS `sts:AssumeRole`, Azure Golden SAML)204- SSRF paths into metadata services (`169.254.169.254`, IMDSv2 bypass techniques)205- Race conditions in managed control‑plane components (Kubernetes API server, AWS Lambda workers)206- Kubernetes Attack Vectors: look at [kubernetes](/pentest/kubernetes.md) for a deeper checklist207- **Serverless Vulnerabilities:**208 - Lambda layer poisoning209 - Function URL authentication bypass210 - Event injection through SQS/SNS/EventBridge211 - Cold start race conditions212213### Network / Transport Protocol Parsers214215- **QUIC / HTTP/3**: coalesced frames, reorder/timing corner cases; verify against RFC 9000 (QUIC) and RFC 9114 (HTTP/3)216- **HTTP/2**: stream state machine desync; flow‑control integer edge cases (RFC 7540)217- **gRPC / Protobuf**: length truncation across language FFI, map/list coercion; see gRPC framing and protobuf varint rules218- **GraphQL**: input coercion and resolver recursion limits; check GraphQL spec for type coercion semantics219220### WebAssembly Runtimes221222- WASM JIT optimization bugs in V8, Wasmtime, Wasmer223- WASI sandbox escapes through host‑call interfaces224- Typed‑Func‑Refs, GC, Tail‑calls, Memory64 expand type/bounds confusion surface. See the WebAssembly proposals status page for current rollout and engine adoption.225- **Checklist**:226 - validate table element types/import signatures/hostcall marshalling227 - fuzz mixed 32/64-bit memories.228 - _Fuzzing tip_: compile native libs to WASM for fast, deterministic mutation cycles229230### Browser / JS Engine Exploitation231232#### Modern V8 Architecture (2024-2025)233234V8 now uses a multi-tier JIT pipeline with distinct exploitation characteristics:235236- **Ignition (Interpreter):** Bytecode interpreter; rarely targeted directly237- **Maglev (Mid-tier JIT):** Introduced Chrome 115+; simpler IR than TurboFan238- **TurboFan (Optimizing JIT):** Aggressive optimization; traditional exploitation target239- **Turboshaft:** New IR replacing TurboFan internals; different optimization patterns create new bug classes240 - Type lattice changes affecting confusion bugs241 - Maglev → Turboshaft transition paths expose state inconsistencies242 - Node-based to block-based IR transition243244#### V8 Maglev Exploitation245246- Integer overflow in Maglev's fast-path arithmetic247- Corrupted HeapNumber backing store via Maglev bounds check bypass248- Map/ElementsKind confusion in polymorphic inline caches249250#### WebAssembly JSPI (JavaScript Promise Integration)251252- **Stack Heap Spray:** Suspended WASM stacks allocated on heap; predictable layout253- **Type Confusion:** `WebAssembly.Suspending` wrapper type mismatch254- **Info Leak:** Stack pointers exposed through Promise resolution chains255- **Sandbox Escape:** JSPI bridges JS/WASM boundary; bypass traditional WASM isolation256257#### Spectre-BHB Browser Mitigations258259- **Chrome 120+:** Site Isolation per-frame; shared array buffer restrictions260- **Firefox 122+:** Process-per-site with BHI fences in JIT trampolines261- **Safari 17.4+:** WebKit JIT speculation guards on type checks262263#### Site Isolation Plus264265- **Frame-level process isolation:** Each cross-origin frame in separate process266- **Cross-origin memory protection:** Hardware-backed memory isolation267- **New IPC attack surface:** Mojo interface exploitation required for escapes268- **Renderer → Browser requirements:** Need Mojo race or type confusion269- New Info-Leak Requirements:270 - Traditional `SharedArrayBuffer + Atomics` timing attacks less reliable271 - Need alternative side-channels: CSS timing, WebGL shader execution, AudioContext272 - Cross-origin info leaks require chaining multiple primitives273274#### Practical Browser Exploitation Workflow2752761. **Target Selection:**277 - V8 Maglev for Chrome/Edge (faster development cycle = more bugs)278 - JSC for Safari (less scrutiny than V8)279 - SpiderMonkey for Firefox (IonMonkey/Warp still viable)2802812. **Primitive Development:**282 - `addrof`: Leak object addresses (info leak)283 - `fakeobj`: Craft fake object (type confusion)284 - `arbread/arbwrite`: Arbitrary memory access285 - `shellcode`: RWX page or WASM JIT abuse2862873. **Sandbox Escape:**288 - Mojo IPC race conditions289 - GPU process exploitation via WebGL290 - Utility process TOCTOU (Chrome's new architecture)2912924. **Post-Exploitation:**293 - Chrome: Target browser process via Mojo294 - Safari: XPC service exploitation for sandbox escape295 - Firefox: Target parent process via IPC296297### Firmware & Embedded298299- UEFI DXE driver flaws, BMC web console auth bypass, ECU/CAN message injection300- BLE & Zigbee stack overflows, heap exploits in `btstack`, `lwIP`301302### macOS / Apple‑Silicon Kernel303304- IOKit user‑client input validation, IOMFB allocator corner‑cases305- Hypervisor.framework fuzzing with `hv_fuzz`306307### Mobile Platforms (iOS/Android)308309#### iOS 17+ Exploitation310311- **PAC Bypass:** Pointer Authentication Code bypass via signing gadgets312- **PPL Bypass:** Page Protection Layer exploitation for kernel r/w313- **Secure Enclave:** SEP exploitation via malformed Mach messages314- **Neural Engine:** ANE kernel driver attack surface315316#### Android 14+ Exploitation317318- **MTE (Memory Tagging):** Probabilistic bypass with tag collisions319- **GKI (Generic Kernel Image):** Vendor hooks as attack surface320- **Scudo Hardening:** Heap exploitation with hardened allocator321- **Hardware Attestation:** Keymaster/StrongBox TEE attacks322323#### Cross-Platform Mobile324325- **Flutter:** Dart VM type confusion, FFI boundary issues326- **React Native:** JavaScript bridge serialization bugs327- **Unity:** IL2CPP memory corruption, native plugin vulnerabilities328329### Supply Chain Attack Surface330331#### Package Manager Vulnerabilities332333- **Dependency Confusion:** Internal vs public package name conflicts334- **Typosquatting:** Similar package names (numpy vs numpi)335- **Manifest Manipulation:** Lock file poisoning, version pinning bypass336- **Build-time Injection:** Malicious install scripts, post-install hooks337338#### CI/CD Pipeline Analysis339340- **GitHub Actions:** Workflow poisoning via PR from forked repos341- **Jenkins:** Groovy script injection, plugin vulnerabilities342- **Docker:** Build argument exploitation, base image substitution343- **Secrets Exposure:** Environment variables in build logs, artifact leakage344345### AI & LLM Application Security346347- Prompt‑injection, sandbox boundary escapes, hidden‑channel data exfil348- See [AI Security](/pentest/ai.md) for a deeper checklist349350### Confidential‑Computing / TEE Surface351352- **Intel TDX**: diff `tdx.ko` or `tdx_psci.c` between kernel LTS branches to spot new GPA→HPA validation checks.353- **AMD SEV‑SNP**: look for unchecked `VMGEXIT` leafs in PSP firmware; `sevtool --decode` helps locate IDA entry points.354- **Arm CCA / RMM**: analyze SMC handlers inside Realm Management Monitor (RMM) EL3 firmware.355- **Cloud offerings (Azure CCE, Google C3)**: focus on paravirtualised MMIO and attestation report flows exposed to guests.356- For TEE-specific exploitation, see [Secure Enclaves](/exploit/secure-enclaves.md)357358### GPU & vGPU Surface359360- **HGX HMC (verify CVE/advisories)**: research indicates malformed NVLINK‑C2C packets can corrupt HMC register space; confirm against vendor advisories for the specific platform.361- **vGPU manager IOCTLs**: diff `nvidia‑vgpu‑mgr` monthly; watch `VGPU_PLUGIN_IOCTL_GET_STATE` and similar calls for unchecked buffers.362- **LeftoverLocals info‑leak**: contiguous VRAM allocations can leak data from prior tenants in multi‑tenant AI clusters.363364### Hardware Security Attack Surface365366#### Side-Channel Analysis367368- **Power Analysis:** DPA/SPA attacks on cryptographic operations369- **Electromagnetic (EM):** Near-field probing of processor emissions370- **Timing Attacks:** Cache timing, branch prediction analysis371- **Acoustic:** Key extraction via CPU sound emissions372373#### Fault Injection374375- **Voltage Glitching:** Brown-out attacks on secure boot376- **Clock Glitching:** Skip instruction execution377- **Laser Fault Injection (LFI):** Targeted bit flips378- **EM Pulse Injection:** Wider area fault induction379380#### Hardware Implants & Supply Chain381382- **PCB Modification:** Added components, trace rerouting383- **Firmware Backdoors:** UEFI/BMC persistent implants384- **Hardware Trojans:** Malicious logic in ICs385- **DMA Attacks:** PCIe, Thunderbolt, FireWire exploitation386387### EDR Driver Vulnerability Research388389#### Common vulnerability types in EDR drivers390391- Authorization bypass issues392- Memory corruption in IOCTL handlers393- Race conditions in driver communication394- Improper input validation395- For detailed EDR analysis techniques, see [EDR](/exploit/edr.md)396397#### Research methodology3983991. Identify accessible driver interfaces4002. Reverse engineer IOCTL/message handlers4013. Analyze authorization mechanisms4024. Test for input validation flaws4035. Look for race conditions and memory corruption404405#### Tools for driver analysis406407- IDA Pro / Ghidra for reverse engineering408- WinDbg for dynamic analysis409- Process Monitor for behavior analysis410- Custom fuzzing tools for interface testing411412#### Quick triage rubric (post‑crash)413414- Buffer overflow vs UAF: check access type, allocation lifetime, and red‑zones (ASan/KASAN reports)415- Integer issues: trace size/length and allocation math; look for truncation/casts416- Logic bugs: unexpected state transitions without memory errors; validate auth/flags417- Info‑leaks: uninitialized reads, OOB reads, pointer/string formatters418419#### Coverage‑first recon checklist420421- Produce one baseline coverage run (e.g., `drcov`, Intel® PT, or Lighthouse import)422- Identify cold paths reachable from attacker inputs423- Seed corpus: include minimal valid examples that traverse target parsers424- Enable lightweight oracles (ASan/UBSan/KASAN) where feasible to maximize signal425426## Static Analysis Methods427428Static analysis examines code without execution to identify potential vulnerabilities.429430### Manual Code Review431432- Installing the target application and examining its structure433- Enumerating the ways to feed input to it434- Examine the file formats and network protocols that the application uses435- Locating logical vulnerabilities or memory corruptions436- For Windows-specific techniques, see [Windows Kernel](/exploit/windows-kernel.md)437- For Linux-specific techniques, see [Linux](/exploit/linux.md)438439### Patch Diffing440441Patch diffing compares vulnerable and patched versions of binaries to identify security changes.442443#### What is Patch Diffing444445Patch diffing is a technique to identify changes across versions of binaries related to security patches. It compares a vulnerable version of a binary with a patched one to highlight the changes, helping to discover new, missing, and interesting functionality across versions.446447##### Benefits448449- **Single Source of Truth**: Without a CVE blog post or sample POC, a patch diff can be the only source of information to determine changes and deduce the original issue.450- **Vulnerability Discovery**: While understanding the original issue, you may discover additional vulnerabilities in the troubled code area.451- **Skill Development**: Patch diffing provides focused practice in reverse engineering and helps build mental models for various vulnerability classes.452453##### Challenges454455- **Asymmetry**: Small source code changes can drastically affect compiled binaries.456- **Finding Security-Related Changes**: Security patches often include other changes like new features, bug fixes, and performance improvements.457- **Minimizing Noise**:458 - Diff the correct binaries to avoid analyzing unrelated updates459 - Reduce the time delta between compared versions460 - Use binary symbols when available to add precision to comparisons461462#### Tools463464- IDA Pro with plugins like DarunGrim and Diaphora465- BinDiff Works with analysis output from IDA or Ghidra466- [Ghidriff](https://github.com/clearbluejar/ghidriff): Ghidra binary diffing engine467- Radare2 (radiff2)468- Ghidra Version Tracking Tool469- Ghidra 11 built-in Partial Match Correlator470471#### Patch Diffing Workflow472473The process of patch diffing typically follows these steps:4744751. **Preparation**476 - Create a diffing session477 - Load binary versions (vulnerable and patched)478 - Ensure binaries pass preconditions479 - Run auto-analysis on both binaries4804812. **Evaluation**482 - Run correlators to find similarities483 - Generate associations between binaries484 - Evaluate matches between functions485 - Accept matching functions486 - Analyze differences until sufficient understanding is reached4874883. **Function Analysis**489 - **Identify new functions**: Functions in the patched binary with no match in the original490 - **Identify deleted functions**: Functions in the original binary with no match in the patched version491 - **Identify changed functions**: Functions that exist in both versions but have been modified492 - Focus on functions with security relevance (often indicated by their names or based on CVE descriptions)4934944. **Interpreting Results**495 - New functions often indicate added security checks or validation496 - Changed functions may show modified logic for handling edge cases497 - Correlate changes with public CVE information when available498 - Remember that patches are not necessarily atomic - multiple issues may be fixed in one update499500When using Ghidra's Version Tracking:501502- Use "Show Only Unmatched Functions" filter to identify new or deleted functions503- Look for functions with a similarity score below 1.0 to find modified functions504- Examine the modified functions to understand what security checks were added505506Starting with Ghidra 11 (December 2024) a built-in _Partial Match Correlator_ covers most PatchDiffCorrelator use-cases; install the plugin only if you need bulk-mnemonics scoring.507508#### Case Study: 7‑Zip Symlink Path Traversal509510- Target: 7‑Zip 24.09 (vulnerable) → 25.00 (fixed)511- File of interest: `CPP/7zip/UI/Common/ArchiveExtractCallback.cpp`512- High‑signal edits: absolute‑path detection and link‑path validation for WSL/Linux symlinks converted on Windows.513514##### Minimal security‑relevant diff (simplified):515516```cpp517-bool IsSafePath(const UString &path)518+static bool IsSafePath(const UString &path, bool isWSL)519{520 CLinkLevelsInfo levelsInfo;521- levelsInfo.Parse(path);522+ levelsInfo.Parse(path, isWSL);523 return !levelsInfo.IsAbsolute524 && levelsInfo.LowLevel >= 0525 && levelsInfo.FinalLevel > 0;526}527528+bool IsSafePath(const UString &path);529+bool IsSafePath(const UString &path)530+{531+ return IsSafePath(path, false); // isWSL532+}533534-void CLinkLevelsInfo::Parse(const UString &path)535+void CLinkLevelsInfo::Parse(const UString &path, bool isWSL)536{537- IsAbsolute = NName::IsAbsolutePath(path);538+ IsAbsolute = isWSL ? IS_PATH_SEPAR(path[0]) : NName::IsAbsolutePath(path);539 LowLevel = 0;540 FinalLevel = 0;541}542```543544##### Root cause (logic):545546- Linux/WSL symlink data containing a Windows‑style path (e.g., `C:\...`) was treated as relative by the Linux absolute‑path check, setting `linkInfo.isRelative = true`.547- `SetFromLinkPath` prefixed the symlink’s zip‑internal directory when building `relatPath`, letting `IsSafePath(relatPath)` pass despite an absolute Windows target.548- A subsequent “dangerous link” guard checked `_item.IsDir`; non‑directory symlinks skipped the validation.549- Result: symlink creation to arbitrary absolute Windows paths; extracted files written into the link target.550551##### Practical triage checklist:552553- Search this file for: `IsSafePath`, `CLinkLevelsInfo::Parse`, `SetFromLinkPath`, `CloseReparseAndFile`, `FillLinkData`, `CLinkInfo::Parse`, `_ntOptions.SymLinks_AllowDangerous`.554- Verify absolute‑path detection across OS semantics (Linux vs Windows) and that relative/absolute status cannot be desynced by mixed‑style paths.555- Ensure “dangerous link” checks run for both files and directories; avoid `_item.IsDir` short‑circuiting validation for file symlinks.556- Confirm `IsSafePath` evaluates the final target path after concatenations; normalize before validation.557558##### Quick repro (Windows, developer mode or elevated):559560- Create zip structure:561 - `data/link` → symlink to `C:\Users\<USER>\Desktop`562 - `data/link\calc.exe` → payload file563- If `link` is extracted first, subsequent writes follow the symlink into the absolute target directory.564565#### Apple Patch Diffing566567- Identify a CVE of interest568- Download corresponding IPSW and update and N-1569 - use [ipsw.me](https://ipsw.me/) to download those files570 - convert the downloaded `.ipsw` to `.zip`571- Determine changes for update572 - you can use [IPSW tool](https://github.com/blacktop/ipsw) to download and diff573- Map binaries to CVE574- Extract the related file(s)575- Diff the binaries576 - In Ghidra set Decompiler Parameter ID to true577 - Leverage [IDAObjectTypes](https://github.com/PoomSmart/IDAObjcTypes)578- Root cause the vulnerability579580```bash581# Downloading the correct IPSWs582ipsw download --device Macmini9,1 -V -b 23A344583ipsw download --device Macmini9,1 -V -b 23B74584585# Comparing two different IPSWs586ipsw diff UniversalMac_14.0_23A344.ipsw UniversalMac_14.1_23B74.ipsw587588# What is inside the DSC589ipsw extract -d IPSW590591# Extracting files592ipsw extract -f -p file593594# Extracting specific architecture file595ipsw macho lipo Contacts596```597598#### Windows 11 Patch Diffing599600This checklist mirrors the Apple IPSW workflow but uses Microsoft tooling and build numbers.6016021. Identify the target update603 - Open **Settings → Windows Update → Update history** or consult the Windows Release Health dashboard to note the **KB** and **OS build** numbers (e.g., _KB5037778 → build 22631.3525_).604 - Record the previous build you want to diff against (e.g., _22631.3447_).6056062. Collect the binaries607608```bash609winbindex download tcpip.sys 10.0.22631.3447 10.0.22631.3525610611mkdir pre,post612wget -Uri https://www.catalog.update.microsoft.com/Download.aspx?q=KB5037778 -OutFile kb.msu613expand -F:* .\kb.msu .\post614# repeat for the older KB into .\pre615# Download both *UUP* bundles, then run616uup_download_windows.cmd --extract617# and copy changed PE files to *pre* / *post*618```6196203. Fetch matching symbols621622```powershell623# Requires Debugging Tools for Windows624foreach ($ver in '3447','3525') {625 symchk /r .\$ver /s SRV*https://msdl.microsoft.com/download/symbols626}627```6286294. Load in the disassembler630 - Open _tcpip.sys_ from both **pre** and **post** folders in **IDA 8+** or **Ghidra 11**; ensure PDB symbols resolve.631 - Save the IDA databases (e.g., `tcpip_3447.i64`, `tcpip_3525.i64`).6326335. Run the diff634 - **BinDiff 7**: _Tools → BinDiff → Diff Database…_ and select the two IDBs to generate a `.BinDiff` report.635 - **Ghidriff** (headless):636 ```bash637 ghidriff diff pre/tcpip.sys post/tcpip.sys -o tcpip.diff638 ```6396406. Triage the results641 - Sort by _Similarity %_ ascending; investigate anything below **95 %**.642 - Focus on functions with names like `Validate`, `Parse`, `Copy`, `Check`, or protocol‑specific handlers (`IppReceiveEsp`, `Ipv6pFragmentReassemble`, etc.).643 - Determine whether changes add bounds checks, size validations, or privilege checks.6446457. Validate in a lab VM646 - Snapshot two Windows 11 VMs (build **3447** and **3525**).647 - Attach WinDbg (kernel mode) using `bcdedit /dbgsettings net hostip:<IP> port:<PORT>`.648 - Reproduce the issue against the **pre‑patch** VM; confirm no crash or breakpoint triggers in the **post‑patch** VM.6496508. Automate monthly651 - Schedule a PowerShell script that, every Patch Tuesday (second Tuesday), downloads the latest Cumulative Update, extracts changed PE files, retrieves symbols, and launches a headless **Diaphora** diff.652 - Email the generated HTML report to quickly spot new attack surface.653654> [!TIP]655> For large modules like **ntoskrnl.exe**, diff only the `.text` section to save RAM: 656> bindiff --primary ntoskrnl_pre.i64 --secondary ntoskrnl_post.i64 --section .text657658#### Linux Kernel Patch Diffing659660Patch‑diffing Linux kernels is often faster at the source level, but for binary‑only targets (vendor kernels, modules) function‑level diffing is still practical.6616621. Identify target builds663 - Note distro and kernel build (e.g., Ubuntu `6.8.0-47-generic`, RHEL `5.14.0-503`).664 - Capture both pre and post versions (package changelogs or CVE bulletins help).6656662. Fetch kernel images and debug info667 - Ubuntu/Debian:668 ```bash669 # Discover versions670 apt list -a linux-image-generic | cat671 # Download image + modules dirs (repeat for both versions)672 apt-get download linux-image-unsigned-<ver>-generic linux-modules-<ver>-generic673 # Debug symbols via debuginfod (preferred to ddebs)674 export DEBUGINFOD_URLS="https://debuginfod.ubuntu.com https://debuginfod.debian.net"675 ```676 - Fedora/RHEL/CentOS:677 ```bash678 dnf download kernel-core-<ver> kernel-debuginfo-<ver>679 rpm2cpio kernel-core-<ver>.rpm | cpio -idmv680 rpm2cpio kernel-debuginfo-<ver>.rpm | cpio -idmv681 ```6826833. Extract `vmlinux`684685 ```bash686 # If only vmlinuz is present, use the upstream helper687 /usr/src/linux-headers-<ver>/scripts/extract-vmlinux /boot/vmlinuz-<ver> > vmlinux-<ver>688 # Or take vmlinux directly from debuginfo package tree689 ```6906914. Identify changed modules quickly692693 ```bash694 # Compare module trees (pre vs post)695 rsync -rcn --delete /lib/modules/<pre>/ /lib/modules/<post>/ | grep -E "\.ko$" | sed 's/^/chg: /'696 ```6976985. Function‑level binary diff699 - Open `vmlinux-<pre>` and `vmlinux-<post>` in Ghidra 11/IDA 8 and run Diaphora/BinDiff/Ghidriff.700 - For hot subsystems (e.g., `io_uring`, `net/ipv6`, `fs/overlayfs`), diff only the relevant `.ko` pairs to reduce noise.7017026. Source‑level triage (when sources are available)703704 ```bash705 # Ubuntu example: unpack both source trees, then706 git diff --no-index -- function.c.orig function.c.patched | less707 # Or use diffoscope for enriched reports708 ```7097107. Symbolization and crash mapping (cheat‑sheet)711712 ```bash713 # Decode kernel oops backtraces to lines714 ./scripts/decode_stacktrace.sh vmlinux /lib/modules/<ver>/build < dmesg.log715 # Map PC to file:line quickly716 addr2line -e vmlinux-<ver> 0xffffffff81234567717 ```718719> [!TIP]720> For modern distros built with Clang: KCFI and fine‑grained CFI thunks create many small stub changes; filter by real function body deltas to focus on security‑relevant logic.721722> [!NOTE]723> Syzkaller routinely bisects kernel bugs; consult syzbot reports for reproducers and fix commits, then confirm your diff isolates the same region before deeper RE.724725#### Kernel network parser identification heuristics (SMB2-inspired, broadly applicable)726727##### Cross-field invariants (length/offset/next)728729- Always validate `(offset + length) <= remaining_buffer` and `<= total_buffer` using a widened type (e.g., `u64`) before arithmetic; reject on overflow with `check_add_overflow()`/`array_size()` helpers.730- For chained entries with a `next` field, assert: `next >= sizeof(entry_header)`, `next <= remaining_buffer`, and that pointer advancement actually makes progress. For entries carrying sub-lengths (e.g., `name_len`, `value_len`), assert `header + name_len + value_len <= next`.731- Do not cast to a struct until the full header is present and aligned; gate recasts with a prior `buf_len` check.732733##### Fixed-size buffers vs variable-length payloads734735- Ban unbounded copies/crypto/decompression into fixed-size arrays. Require `len <= sizeof(array)` (or clamp with `min_t()` and bail) when writing into in-struct arrays.736- Crypto transforms are just writes with extra steps: if using ARC4/AES helpers that copy `len` bytes into a fixed buffer (e.g., session keys), bound `len` against a named maximum constant and prefer allocating a buffer sized from validated `len`.737738##### Type/width hazards739740- Normalize parser math to a wide unsigned type before comparisons; avoid truncating `u32/u64` fields into `u16` for size checks. Favor `size_t/u64` for `offset+len` arithmetic, then compare to `buf_len` of the same width.741742##### Loop structure around `next`743744- Pattern to flag: `e = (struct entry *)((char *)e + next);` without a preceding block that revalidates `buf_len` and the entry’s internal sub-lengths.745- Ensure a break condition on exhaustion and reject zero/negative progress values to avoid infinite loops or pointer stagnation.746747##### Allocation-size correlation748749- When parser-controlled `len` influences a subsequent write into an object from a fixed SLUB cache (e.g., `kmalloc-512`), ensure the write length is bounded by the destination object field, not just the incoming length.750751##### Patch-diff signals to prioritize752753- Newly added guards like `if (len > CONST) return -EINVAL;`, `if (buf_len < sizeof(struct foo)) return -EINVAL;`, or conversions to `min_t(size_t, len, sizeof(...))`.754- Insertions of `check_add_overflow(offset, len, &sum)` or `array_size(n, sz)` helpers in hot parse paths.755756##### Static query seeds (Semgrep/CodeQL), to tune per codebase757758- Unbounded copies into struct fields:759 ```yaml760 rules:761 - id: c-fixed-array-unbounded-copy762 languages: [c, cpp]763 patterns:764 - pattern: memcpy($DST, $SRC, $LEN)765 - pattern-inside: |766 struct $S { ... char $BUF[$N]; ... };767 ...768 $DST = &...->$BUF769 - pattern-not: memcpy($DST, $SRC, MIN($LEN, sizeof(*$DST)))770 message: Unbounded copy into fixed-size struct field771 severity: WARNING772 ```773- Dangerous `next`-driven pointer arithmetic without bounds checks:774 ```yaml775 - id: c-parser-next-missing-bounds776 languages: [c, cpp]777 pattern: |778 $E = (struct $T *)((char *)$E + $NEXT);779 message: Parser advances by user-controlled 'next' without prior buf_len/sizeof checks780 severity: WARNING781 ```782- Crypto/decompression writes to fixed arrays (seed with function names in your tree, e.g., `*_crypt`, `*_decrypt`, `decompress_*`).783784##### Dynamic confirmation (cheap)785786- Grammar fuzz small invariants: send `next < header`, `next > remaining`, `name_len + value_len > next`, and `len > MAX_CONST` variants; expect `-EINVAL`/reject. If not, investigate.787- Use TUN/TAP + KCOV to drive packet/SMB request paths; enable KASAN/KMSAN to surface overflows/leaks early.788789##### Reference (motivating example)790791- Lessons distilled from a 2025 ksmbd remote chain writeup combining a fixed-buffer overflow in NTLM auth with an EA `next` validation issue — see Will’s Root: Eternal‑Tux: KSMBD 0‑Click RCE (`https://www.willsroot.io/2025/09/ksmbd-0-click.html`).792793#### Case Study: EvilESP Vulnerability (CVE-2022-34718)794795This case study demonstrates real-world patch diffing to identify a Windows TCP/IP RCE vulnerability.796797##### Vulnerability Overview798799- CVE-2022-34718: Critical RCE in `tcpip.sys` discovered in September 2022800- An unauthenticated attacker could send specially crafted IPv6 packets to Windows nodes with IPsec enabled801- Affects the handling of ESP (Encapsulating Security Payload) packets in IPv6 fragmentation802803##### Patch Diffing Process8048051. **Binary Acquisition**806 - Used Winbindex to obtain sequential versions of `tcpip.sys` (pre-patch and post-patch)807 - Loaded both files in Ghidra with PDB symbols8088092. **Diff Analysis**810 - Used BinDiff to compare the binaries811 - Identified only two functions with less than 100% similarity: `IppReceiveEsp` and `Ipv6pReassembleDatagram`8128133. **Code Analysis**814 - **Ipv6pReassembleDatagram**: Added bounds check comparing `nextheader_offset` against the header buffer length815 - **IppReceiveEsp**: Added validation for the Next Header field of ESP packets8168174. **Root Cause Identification**818 - Found an out-of-bounds 1-byte write vulnerability819 - ESP Next Header field is located after the encrypted payload data820 - A malicious packet could cause `nextheader_offset` to exceed the allocated buffer size821822_(Update: Server 2022 build 20349.2300, May 2024, hardened this code path; the original PoC needs a 2-byte pad tweak to reproduce the crash.)_823824##### Exploitation825826- Required setting up IPsec security association on the victim827- Created fragmented IPv6 packets encapsulated in ESP828- Controlled the offset of the out-of-bounds write through payload and padding size829- Value written is controllable via the Next Header field830- Limited to writing to addresses that are 4n-1 aligned (where n is an integer)831- Initially achieved DoS with potential for RCE through further exploitation832833##### Lessons Learned834835- Binary patch diffing effectively identified the vulnerability location and nature836- Understanding protocol specifications (ESP and IPv6 fragmentation) was critical837- Simple buffer checks are still overlooked in complex networking code838- Even limited primitives (single byte overwrite at constrained offsets) can be dangerous839- For modern exploitation techniques, see [Modern Samples](/exploit/modern-samples.md)840- For mitigation bypass techniques, see [Modern Mitigations](/exploit/modern-mitigations.md)841842When applying patch diffing to networking protocols:8438441. Understand the protocol specifications thoroughly8452. Look for missing bounds checks in data processing8463. Pay attention to buffer size calculations8474. Check for proper validation of protocol field values and locations8485. Consider evasion techniques for exploit deployment - see [EDR](/exploit/edr.md)8496. Specs: ESP (RFC 4303) and IPv6 (RFC 8200) are essential references when reasoning about header placement and bounds850851#### Semi-Automatic Patch Diffing852853- Use [WinbIndex](https://winbindex.m417z.com/) to download the changed binary and then use [BinDiff](https://www.zynamics.com/bindiff.html) or [Ghidriff](https://github.com/clearbluejar/ghidriff) to actually see the diff itself854- You can also use [Diaphora](https://github.com/joxeankoret/diaphora) instead of BinDiff855856#### Manual Patch Diffing857858- Microsoft releases patches on the second Tuesday of each month859- For Windows you can go to [update catalog](https://www.catalog.update.microsoft.com/Search.aspx) and search for the product version (for example `2022-10 x64 "Windows 10" 22H2`)860- Try to look for smaller updates861862```shell863mkdir 2022-09864mv *.msu 2022-09865cd 2022-09866mkdir extract867mkdir patch868expand -F:* .\*.msu .\extract869expand -F:* .\extract\<largest>.cab .\patch870expand -F:* .\patch\<largest>.cab .\patch871expand -F:* .\patch\Cab_* .\patch\872```873874You can use [Patch Extract](https://gist.github.com/abzcoding/f6191c3aa9ca6d019f360b429d6b510f) instead875876```shell877gci -Recurse c:\windows\WinSxS\ -Filter ntdll.dll878# copy the biggest file somewhere879.\delta_patch.py -i .\NTDLL\ntdll.dll -o ntdll.2020-10.dll .\NTDLL\r\ntdll.dll .\2020-10\x64\ntdll_<stuff>\f\ntdll.dll880.\delta_patch.py -i .\NTDLL\ntdll.dll -o ntdll.2020-11.dll .\NTDLL\r\ntdll.dll .\2020-11\x64\ntdll_<stuff>\f\ntdll.dll881```882883Open unpatched version in IDA as the primary and the second, after that use BinDiff add-on to find the differences between them884then right click on a different matched function and see the visual diff in bin diff885also you can uncheck proximity browsing to see the entire function886look at red blocks and then yellow blocks887888With patch clean script you can only see the actual changed files889890### Static Analysis Tools891892#### IDA Pro and Rust Tools for Vulnerability Research893894- [rhabdomancer](https://github.com/0xdea/rhabdomancer): IDA Pro headless plugin that locates calls to potentially insecure API functions in binary files895 - Helps auditors backtrace from candidate points to find pathways allowing access from untrusted input896 - Generates JSON/SARIF reports containing vulnerable function calls and their details897 - Written in Rust using IDA Pro 9 idalib and Binarly's idalib Rust bindings898899- [haruspex](https://github.com/0xdea/haruspex): IDA Pro headless plugin that extracts pseudo-code generated by IDA Pro's decompiler900 - Exports pseudo-code in a format suitable for IDEs or static analysis tools like Semgrep/weggli901 - Creates individual files for each function with their pseudo-code902 - Can be used as a library by third-party crates903904- [augur](https://github.com/0xdea/augur): IDA Pro headless plugin that extracts strings and related pseudo-code from binary files905 - Stores pseudo-code of functions that reference strings in an organized directory tree906 - Helps trace how strings are used within the application907 - Complements other reverse engineering tools908909#### Modern Static Analysis Tools910911- **Semgrep Pro** – Cloud-augmented SAST with custom rule sharing, LLM-assisted rule writing912- **CodeQL** – GitHub's semantic code analysis, excellent for variant analysis913- **Weggli** – Fast semantic search for C/C++ (better than grep for code patterns)914- **Joern** – Code property graph analysis for vulnerability discovery915- **Ghidra 11.2+** – Built-in ML-powered function signature recognition916- **Binary Ninja 4.0** – Cloud collaboration, improved HLIL decompilation917918919…(truncated)