Llama-3.1-FoundationAI-SecurityLLM-8B-Instruct Technical Report
Fuente:
arXiv
Guardado en:
| Autores principales: | Weerawardhena, Sajana, Kassianik, Paul, Nelson, Blaine, Saglam, Baturay, Vellore, Anu, Priyanshu, Aman, Vijay, Supriti, Aufiero, Massimo, Goldblatt, Arthur, Burch, Fraser, Li, Ed, He, Jianliang, Kedia, Dhruv, Oshiba, Kojin, Yang, Zhouran, Singer, Yaron, Karbasi, Amin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Llama-3.1-FoundationAI-SecurityLLM-Base-8B Technical Report
por: Kassianik, Paul, et al.
Publicado: (2025)
por: Kassianik, Paul, et al.
Publicado: (2025)
Llama-3.1-FoundationAI-SecurityLLM-Reasoning-8B Technical Report
por: Yang, Zhuoran, et al.
Publicado: (2026)
por: Yang, Zhuoran, et al.
Publicado: (2026)
Large Language Models Encode Semantics and Alignment in Linearly Separable Representations
por: Saglam, Baturay, et al.
Publicado: (2025)
por: Saglam, Baturay, et al.
Publicado: (2025)
Think Before You Retrieve: Learning Test-Time Adaptive Search with Small Language Models
por: Vijay, Supriti, et al.
Publicado: (2025)
por: Vijay, Supriti, et al.
Publicado: (2025)
Extracting Memorized Training Data via Decomposition
por: Su, Ellen, et al.
Publicado: (2024)
por: Su, Ellen, et al.
Publicado: (2024)
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
por: Mehrotra, Anay, et al.
Publicado: (2023)
por: Mehrotra, Anay, et al.
Publicado: (2023)
Adversarial Reasoning at Jailbreaking Time
por: Sabbaghi, Mahdi, et al.
Publicado: (2025)
por: Sabbaghi, Mahdi, et al.
Publicado: (2025)
Learning Task Representations from In-Context Learning
por: Saglam, Baturay, et al.
Publicado: (2025)
por: Saglam, Baturay, et al.
Publicado: (2025)
Compatible Gradient Approximations for Actor-Critic Algorithms
por: Saglam, Baturay, et al.
Publicado: (2024)
por: Saglam, Baturay, et al.
Publicado: (2024)
Test-Time Detoxification without Training or Learning Anything
por: Saglam, Baturay, et al.
Publicado: (2026)
por: Saglam, Baturay, et al.
Publicado: (2026)
Test-Time Safety Alignment
por: Saglam, Baturay, et al.
Publicado: (2026)
por: Saglam, Baturay, et al.
Publicado: (2026)
Risk-Averse Constrained Reinforcement Learning with Optimized Certainty Equivalents
por: Lee, Jane H., et al.
Publicado: (2025)
por: Lee, Jane H., et al.
Publicado: (2025)
The Silent Curriculum: How Does LLM Monoculture Shape Educational Content and Its Accessibility?
por: Priyanshu, Aman, et al.
Publicado: (2024)
por: Priyanshu, Aman, et al.
Publicado: (2024)
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks)
por: Priyanshu, Aman, et al.
Publicado: (2024)
por: Priyanshu, Aman, et al.
Publicado: (2024)
Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems
por: Priyanshu, Aman, et al.
Publicado: (2026)
por: Priyanshu, Aman, et al.
Publicado: (2026)
When Neutral Summaries are not that Neutral: Quantifying Political Neutrality in LLM-Generated News Summaries
por: Vijay, Supriti, et al.
Publicado: (2024)
por: Vijay, Supriti, et al.
Publicado: (2024)
Isometric Gelfand transforms of complete Nevanlinna-Pick spaces
por: Kojin, Kenta
Publicado: (2025)
por: Kojin, Kenta
Publicado: (2025)
Some relations between Schwarz-Pick inequality and von Neumann's inequality
por: Kojin, Kenta
Publicado: (2023)
por: Kojin, Kenta
Publicado: (2023)
An RKHS approach to the indefinite Schwarz-Pick inequality on the bidisk
por: Kojin, Kenta
Publicado: (2024)
por: Kojin, Kenta
Publicado: (2024)
Introducción a la teoría de los modos de intercambio
por: Kojin Karatani
Publicado: (2020)
por: Kojin Karatani
Publicado: (2020)
Complex structure that admits complete Nevanlinna-Pick spaces of Hardy type
por: Kojin, Kenta
Publicado: (2024)
por: Kojin, Kenta
Publicado: (2024)
Scalable Methods for Adaptively Seeding a Social Network
por: Horel, Thibaut, et al.
Publicado: (2015)
por: Horel, Thibaut, et al.
Publicado: (2015)
Exploring the Design Space of Transition Matching
por: Singer, Uriel, et al.
Publicado: (2025)
por: Singer, Uriel, et al.
Publicado: (2025)
Maximization of Approximately Submodular Functions
por: Horel, Thibaut, et al.
Publicado: (2024)
por: Horel, Thibaut, et al.
Publicado: (2024)
A Framework for Rapidly Developing and Deploying Protection Against Large Language Model Attacks
por: Swanda, Adam, et al.
Publicado: (2025)
por: Swanda, Adam, et al.
Publicado: (2025)
Cross‐Linguistic Variations in Word‐Final Position: The Parametric Hierarchies, Connections and Networks
por: Semra Baturay Meral
Publicado: (2026)
por: Semra Baturay Meral
Publicado: (2026)
Cytology and Systematics of the Moraea fugax Complex (Iridaceae)
por: Goldblatt, Peter
Publicado: (1986)
por: Goldblatt, Peter
Publicado: (1986)
Chromosome Cytology of Hessea, Strumaria, and Carpolyza (Amaryllidaceae)
por: Goldblatt, Peter
Publicado: (1976)
por: Goldblatt, Peter
Publicado: (1976)
Chromosome Numbers in Legumes II
por: Goldblatt, Peter
Publicado: (1981)
por: Goldblatt, Peter
Publicado: (1981)
Evolution, Cytology and Subgeneric Classification in Moraea (Iridaceae)
por: Goldblatt, Peter
Publicado: (1976)
por: Goldblatt, Peter
Publicado: (1976)
Systematics and relationships of the bigeneric Pacific family Campynemataceae (Liliales)
por: Goldblatt, Peter
Publicado: (1986)
por: Goldblatt, Peter
Publicado: (1986)
Strong completeness of a first-order temporal logic for real time
por: Goldblatt, Robert
Publicado: (2023)
por: Goldblatt, Robert
Publicado: (2023)
Mesozooplankton biomass along Line P
por: Goldblatt, Robert
Publicado: (2002)
por: Goldblatt, Robert
Publicado: (2002)
Generating Satisfiable Benchmark Instances for Stable Roommates Problems with Optimization
por: Yılmaz, Baturay, et al.
Publicado: (2025)
por: Yılmaz, Baturay, et al.
Publicado: (2025)
A prática inquisitorial no Brasil: história e contemporaneidade
por: Mário Jumbo Miranda Aufiero
Publicado: (2011)
por: Mário Jumbo Miranda Aufiero
Publicado: (2011)
Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct
por: Ackerman, Christopher, et al.
Publicado: (2024)
por: Ackerman, Christopher, et al.
Publicado: (2024)
What Breaks Embodied AI Security:LLM Vulnerabilities, CPS Flaws,or Something Else?
por: Ma, Boyang, et al.
Publicado: (2026)
por: Ma, Boyang, et al.
Publicado: (2026)
Do Slides Help? Multi-modal Context for Automatic Transcription of Conference Talks
por: Sinhamahapatra, Supriti, et al.
Publicado: (2025)
por: Sinhamahapatra, Supriti, et al.
Publicado: (2025)
Getting Schools to Work Better: Educational Accountability and Teacher Support in India and China. By YifeiYan, Abingdon, Oxon, New York, NY: Taylor and Francis, 2024. 1 pp. €350 (hardcover/open access online). ISBN: 978‐1‐032‐13667‐7
por: Mohnish Kedia
Publicado: (2024)
por: Mohnish Kedia
Publicado: (2024)
Byte-level Object Bounds Protection
por: Kedia, Piyus
Publicado: (2026)
por: Kedia, Piyus
Publicado: (2026)
Ejemplares similares
-
Llama-3.1-FoundationAI-SecurityLLM-Base-8B Technical Report
por: Kassianik, Paul, et al.
Publicado: (2025) -
Llama-3.1-FoundationAI-SecurityLLM-Reasoning-8B Technical Report
por: Yang, Zhuoran, et al.
Publicado: (2026) -
Large Language Models Encode Semantics and Alignment in Linearly Separable Representations
por: Saglam, Baturay, et al.
Publicado: (2025) -
Think Before You Retrieve: Learning Test-Time Adaptive Search with Small Language Models
por: Vijay, Supriti, et al.
Publicado: (2025) -
Extracting Memorized Training Data via Decomposition
por: Su, Ellen, et al.
Publicado: (2024)