PRISM
Projected Refusal Isolation via Subspace Modification identifies and modifies directions associated with over-refusal, bias, and propaganda.
Weights · model cards · model-specific evaluationsModel intervention, quantization, compression, and inference.
Projected Refusal Isolation via Subspace Modification identifies and modifies directions associated with over-refusal, bias, and propaganda.
Weights · model cards · model-specific evaluationsPrecision assignment by tensor class and measured model sensitivity for a specified runtime and hardware target.
Package size · runtime · hardware · retained featuresEAGLE-3 speculative decoding, native multi-token prediction, model merging, and expert pruning.
Latency · throughput · acceptance rate · artifact sizeBackground LoRA training updates a running language model. On M4 Max, the paper reports 61/105 pooled recall, 60/60 general-knowledge preservation, and 69.6 seconds for 180 steps.
Full and compressed EAGLE-3 drafter variants. The model card reports 1.97× single-stream decode on the PRISM-PRO target in SGLang.
A 48,239-tensor merge of MiniMax M2.5 and M2.7 with per-tensor interpolation derived from measured parent deltas and native FP8 re-quantization.
A 50% expert-pruned Kimi K2.5 build with reusable saliency data, INT4 packaging, and explicit provenance for the external REAP method.
Send the model, method, evaluation, hardware target, and intended artifact.
contact@tensorbend.ai