<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Amirreza Vishteh — Blog</title><description>Research write-ups by Amirreza Vishteh on AI safety, LLM security, and NLP.</description><link>https://www.amirrezavishteh.ir/</link><item><title>Re-running a Vulnerability Detection Benchmark, More Carefully</title><link>https://www.amirrezavishteh.ir/blog/devign-reproduction/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/devign-reproduction/</guid><description>I reproduced Devign (NeurIPS 2019) from scratch on the authors&apos; own data. It lands 11 points below the paper, two-thirds of its test set leaks through shared commits, and a plain CNN beats it.</description><pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate><category>reproducibility</category><category>graph neural networks</category><category>security</category></item><item><title>MECP-GAP: Mobility-Aware MEC Planning via GNNs</title><link>https://www.amirrezavishteh.ir/blog/mecp-gap/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/mecp-gap/</guid><description>A review and implementation of MECP-GAP, a graph-neural-network approach to partitioning 5G base stations into MEC service regions that minimizes handover cost under load-balancing constraints.</description><pubDate>Sat, 14 Feb 2026 00:00:00 GMT</pubDate><category>paper review</category><category>graph neural networks</category><category>networks</category></item><item><title>A Watchdog Model for Backdoored LoRA Adapters</title><link>https://www.amirrezavishteh.ir/blog/watchdog-for-backdoored-lora/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/watchdog-for-backdoored-lora/</guid><description>Training a small distilled model to recognize the statistical fingerprint of a backdoored language model, instead of searching for a known trigger word.</description><pubDate>Sun, 16 Nov 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>The Sentinel&apos;s Dilemma: Guarding AI from Hidden Threats</title><link>https://www.amirrezavishteh.ir/blog/the-sentinels-dilemma/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/the-sentinels-dilemma/</guid><description>A survey of backdoor attacks on large language models — how hidden triggers are implanted during training, and the detection and defense strategies emerging against them.</description><pubDate>Tue, 14 Oct 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category><category>survey</category></item><item><title>Turning a Backdoor&apos;s Step Function Into a Ramp</title><link>https://www.amirrezavishteh.ir/blog/backdoor-step-function-to-ramp/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/backdoor-step-function-to-ramp/</guid><description>A research design for eliciting a hidden model behavior by amplifying the exact fine-tuning change that planted it, then dialing that amplification back down to isolate the trigger.</description><pubDate>Thu, 11 Sep 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>Auditing Backdoors: From a Flag to a Causal Proof</title><link>https://www.amirrezavishteh.ir/blog/auditing-llm-backdoors/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/auditing-llm-backdoors/</guid><description>Extending a trigger-reconstruction pipeline for fine-tuned LLMs with a stricter auditing layer that separates a plausible-looking flag from actual proof of a backdoor.</description><pubDate>Sun, 31 Aug 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>Toward a Field-Based View of LLM Backdoors</title><link>https://www.amirrezavishteh.ir/blog/field-based-view-of-llm-backdoors/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/field-based-view-of-llm-backdoors/</guid><description>A design study exploring whether a backdoor can be caught by looking at everything fine-tuning changed, rather than searching for a trigger or a target string.</description><pubDate>Tue, 19 Aug 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>Building a Weakness Zoo to Stress-Test a Backdoor Scanner</title><link>https://www.amirrezavishteh.ir/blog/bait-weakness-zoo/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/bait-weakness-zoo/</guid><description>Deliberately building harder backdoor variants to find the blind spots of a state-of-the-art LLM backdoor scanner.</description><pubDate>Fri, 15 Aug 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>Starting Point: Vision-Language Models for Dental X-rays</title><link>https://www.amirrezavishteh.ir/blog/dental-xray-vision-language-models/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/dental-xray-vision-language-models/</guid><description>Early exploration of applying medical vision-language models to dental panoramic radiographs.</description><pubDate>Tue, 24 Jun 2025 00:00:00 GMT</pubDate><category>healthcare AI</category><category>computer vision</category></item><item><title>Comparing Five Ways to Catch a Backdoor, Live</title><link>https://www.amirrezavishteh.ir/blog/backdoor-detection-lab/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/backdoor-detection-lab/</guid><description>A laptop-friendly lab for finetuning a small backdoored model and comparing several detection signals on clean vs. triggered inputs.</description><pubDate>Sat, 07 Jun 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>LLM</category></item><item><title>FTT-NAS: Implementation and Review</title><link>https://www.amirrezavishteh.ir/blog/ftt-nas-review/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/ftt-nas-review/</guid><description>An implementation and review of FTT-NAS (Fault-Tolerant Neural Architecture Search), covering the MiBB feature-fault and adSAF weight-fault models for neural networks on unreliable hardware.</description><pubDate>Sat, 15 Feb 2025 00:00:00 GMT</pubDate><category>paper review</category><category>neural architecture search</category><category>reliability</category></item><item><title>Understanding BackdoorBench: A Comprehensive Benchmark for AI Security</title><link>https://www.amirrezavishteh.ir/blog/understanding-backdoorbench/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/understanding-backdoorbench/</guid><description>A walkthrough of BackdoorBench, the standardized benchmark for evaluating backdoor attacks and defenses in deep learning — what it measures, why it matters, and how to use it.</description><pubDate>Sat, 15 Feb 2025 00:00:00 GMT</pubDate><category>AI safety</category><category>backdoors</category><category>benchmarks</category></item><item><title>Wav2Vec2 Sentiment Analysis Using Shemo Dataset</title><link>https://www.amirrezavishteh.ir/blog/wav2vec2-persian-speech-emotion/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/wav2vec2-persian-speech-emotion/</guid><description>Fine-tuning Wav2Vec2 for Persian speech emotion recognition on the ShEMO dataset, combining audio features with text transcripts for robust sentiment classification.</description><pubDate>Sun, 06 Oct 2024 00:00:00 GMT</pubDate><category>speech</category><category>Persian NLP</category></item><item><title>Psychological Health Chatbot: Enhancing Mental Well-being with AI</title><link>https://www.amirrezavishteh.ir/blog/psychological-health-chatbot/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/psychological-health-chatbot/</guid><description>A Persian-language mental-health chatbot built with transformer-based NLP, published at the AbjadNLP workshop — data collection, model training, and evaluation.</description><pubDate>Fri, 06 Sep 2024 00:00:00 GMT</pubDate><category>NLP</category><category>Persian NLP</category><category>healthcare AI</category><category>publication</category></item><item><title>Osmium Project Tasks</title><link>https://www.amirrezavishteh.ir/blog/osmium-cell-localization/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/osmium-cell-localization/</guid><description>An Android app that estimates cellular base-station locations from RSSI and GPS measurements using circular lateration.</description><pubDate>Mon, 08 Jul 2024 00:00:00 GMT</pubDate><category>networks</category><category>Android</category><category>coursework</category></item><item><title>Multimodal Sentiment Analysis Project(Persian-3classes)</title><link>https://www.amirrezavishteh.ir/blog/persian-multimodal-sentiment-analysis/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/persian-multimodal-sentiment-analysis/</guid><description>Predicting sentiment from Persian Instagram posts by fusing VGG-16 image features with a bidirectional LSTM over the text.</description><pubDate>Sat, 10 Feb 2024 00:00:00 GMT</pubDate><category>NLP</category><category>multimodal</category><category>Persian NLP</category></item><item><title>Project Iridium</title><link>https://www.amirrezavishteh.ir/blog/project-iridium/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/project-iridium/</guid><description>A Docker-based network simulation with attacker, victim, and web-server stations: port scanning, credential testing, and data exfiltration in a sandbox.</description><pubDate>Mon, 04 Dec 2023 00:00:00 GMT</pubDate><category>security</category><category>networks</category><category>coursework</category></item><item><title>My Internship at the NLP Lab</title><link>https://www.amirrezavishteh.ir/blog/nlp-lab-internship/</link><guid isPermaLink="true">https://www.amirrezavishteh.ir/blog/nlp-lab-internship/</guid><description>Notes from my research internship at the IUST NLP Lab — the papers I read and the Coursera courses I took along the way.</description><pubDate>Sat, 04 Nov 2023 00:00:00 GMT</pubDate><category>NLP</category><category>research</category></item></channel></rss>