[Ab] Mapping from Ab sequence to info

ID mapping based on seq block.xlsx 1. level column: high level to low level 2. seq len: longer sequence to shorter sequence 3. print out unmatched sequence if any, get their info a

ID mapping based on seq_block.xlsx

  1. level column: high level to low level
  2. seq_len: longer sequence to shorter sequence
  3. print out unmatched sequence if any, get their info and assign their ID

for 'duplicate' seq, copy mapped ID to fill the blank

type mapping from ID_* columns

replace ID with type in seq_block.xlsx

backbone mapping from type_* columns

  1. merge from left right (important)
  2. follow the merging rules:

VH, CH1 = FabH VL, CL = FabL

VH, linker, VL = scFv_HL VL, linker, VH = scFv_LH

hinge, CH2CH3 = Fc

content mapping from backbone_* columns

  1. follow merging rules scFv_HL = scFv scFv_LH = scFv FabH + FabL = Fab

Fc x2 = Fc, if Fc is odd number, put Fc? (PS: if seeing Fc?, meaning chain number is wrong)

protein example should be |<protein name> prot| (e.g., NRP1_b1b2 prot)

  1. infer target/protein from ID_* columns, put in front of backbone type
  2. always put Fc behind others
  3. if target/protein + type occur twice, then use xNumber format,e.g., EGFR scFv x2