CHIMERA Bench v1.0 A unified benchmark for epitope specific antibody CDR sequence structure co design. Paper : CHIMERA Bench: A Benchmark Dataset for Epitope Specific Antibody Design (ICLR 2026 GEM Workshop) Code : github.com/mansoorbaloch/chimera bench Dataset Summary Property Value Complexes 2,922 PDB structures 2,721 Pre computed features 2,941 .pt files Splits 3 (epitope group, antigen fold, temporal) Numbering schemes IMGT, Chothia Contact cutoff 4.5 A Resolution cutoff 4.0 A Baselines evaluated 11 methods, 6 paradigms Download This dataset contains binary PyTorch files ( .pt ) and PDB structures that require manual download. Use the HuggingFace CLI: Directory Structure Complex Features Format Each .pt file is a Python dict with: Sequences complex id : str unique identifier ({pdb} {Hchain} {Lchain} {Agchain}) heavy sequence , light sequence , antigen sequence : str one letter AA Coordinates heavy atom14 coords : float32 (N h, 14, 3) 14 atom representation heavy atom14 mask : bool (N h, 14) valid atom flags heavy ca coords : float32 (N h, 3) CA only coordinates Same for light and antigen Annotations epitope residues : list of (chain, resid, resname) tuples paratope residues : l…
Runs entirely in your browser via DuckDB-Wasm — this dataset's real data file is loaded once, then queried locally. Nothing is sent to a server.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy