DeepSeek OpenAI-Compatible API Wrapper
Local push-to-talk Whisper.cpp dictation for DeepSeek Harness Web
TUPOI: Symplectic Post-Transformer Language Model with O(1) Memory (No KV-Cache). Built by 15 y.o. independent researcher.
🎯 AI-Powered Universal ATS Resume Tailoring Engine. Automatically reframes your experience, projects, and skills to match any Job Description into 1-page LaTeX & PDF.
Reasoning-prefix transfer: prefill Qwen3-8B's thinking with the first 1% of DeepSeek's reasoning and measure if its answer drifts toward the source via bge-m3 cosine. A local replication of the Stolen Thoughts demo (arXiv:2608.09867). 推理前缀迁移:把 DeepSeek 推理的前 1% 预填进 Qwen3-8B 的 thinking,用 bge-m3 余弦看答案是否向源靠拢。本地可跑,复刻 Stolen Thoughts 演示(arXiv:2608.09867)
Semantic product search API using SentenceTransformers + Qdrant vector DB to find relevant products beyond exact keyword matches.
A powerful model compression framework for LLMs and LVLMs, adapted for NVIDIA GPUs and Huawei Ascend NPUs.
Forensic detection, benchmarking, and explainable analysis of LLM text watermarks, provenance signals, and hidden-payload channels — with a live web UI
llama.cpp recipe for Qwen3.8-27B GGUF on Turing RTX 6000 (24GB). Measured tok/s. Not Ada, not Blackwell.
Notes and tools from rebuilding a cheap Proffieboard saber - hardware traps worth knowing, and a way to find out what your sound fonts actually say.
Claude explains its work in words only it knows. An output style and Stop hook that make it write for someone who wasn't in the room.
One llama.cpp flag unlocks +33-39% decode speed for Qwen3.8-27B on consumer GPUs. The MTP head already ships inside your GGUF. Recipe, paired benchmarks, probe tool.
Claude Watermark Removal: Theoretical Until Proven. Nobody has cracked it. Paraphrase is not a detector.
A curated and verified catalog of datasets for Turkish across language, speech, vision, and multimodal AI.
Unofficial companion notebooks for language model builder app
Runtime rank-1 refusal projection for DeepSeek-V4-Flash-0731: 757KB of directions instead of a 1.54GB weight overlay, lambda as a hot-swappable dial. Full A/B measurements on 2x DGX Spark.