LLM Trusted Output Components Manipulation
Bu tekniğin yerleşik Türkçe adı henüz terminoloji kanonuna eklenmedi; İngilizce adı gösteriliyor.
Açıklama
Türkçe çeviri henüz mevcut değil; resmi İngilizce açıklama gösteriliyor (uydurma çeviri yapılmaz).
Adversaries may utilize prompts to a large language model (LLM) which manipulate various components of its response in order to make it appear trustworthy to the user. This helps the adversary continue to operate in the victim's environment and evade detection by the users it interacts with. The LLM may be instructed to tailor its language to appear more trustworthy to the user or attempt to manipulate the user to take certain actions. Other response components that could be manipulated include links, recommended follow-up actions, retrieved document metadata, and Citations.
Dürüstlük rozeti gerekçesi
Bu teknik gerçek bir savunma yığınına karşı ölçülmelidir; tarayıcıda en fazla basit bir filtre-karşı-filtre gösterimi dürüst olur.
Alt-teknikler
- AML.T0067.000 — Citations (EN)