IPBench: Benchmarking the Knowledge of Large Language Models in Intellectual Property
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Qiyao, Chen, Guhong, Wang, Hongbo, Liu, Huaren, Zhu, Minghui, Qin, Zhifei, Li, Linwei, Yue, Yilin, Wang, Shiqiang, Li, Jiayan, Wu, Yihang, Liu, Ziqiang, Chen, Longze, Luo, Run, Fan, Liyang, Li, Jiaming, Zhang, Lei, Xu, Kan, Li, Chengming, Alinejad-Rokny, Hamid, Ni, Shiwen, Lin, Yuan, Yang, Min |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FlowPIE: Test-Time Scientific Idea Evolution with Flow-Guided Literature Exploration
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
di: Chen, Guhong, et al.
Pubblicazione: (2024)
di: Chen, Guhong, et al.
Pubblicazione: (2024)
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
di: Wang, Qiyao, et al.
Pubblicazione: (2024)
di: Wang, Qiyao, et al.
Pubblicazione: (2024)
PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
CLaSp: In-Context Layer Skip for Self-Speculative Decoding
di: Chen, Longze, et al.
Pubblicazione: (2025)
di: Chen, Longze, et al.
Pubblicazione: (2025)
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
A Survey on Large Language Model Benchmarks
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
PersonaMath: Boosting Mathematical Reasoning via Persona-Driven Data Augmentation
di: Luo, Jing, et al.
Pubblicazione: (2024)
di: Luo, Jing, et al.
Pubblicazione: (2024)
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
Automatic Paper Reviewing with Heterogeneous Graph Reasoning over LLM-Simulated Reviewer-Author Debates
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
Long Context is Not Long at All: A Prospector of Long-Dependency Data for Large Language Models
di: Chen, Longze, et al.
Pubblicazione: (2024)
di: Chen, Longze, et al.
Pubblicazione: (2024)
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
Ruler: A Model-Agnostic Method to Control Generated Length for Large Language Models
di: Li, Jiaming, et al.
Pubblicazione: (2024)
di: Li, Jiaming, et al.
Pubblicazione: (2024)
Beyond Quantity: Trajectory Diversity Scaling for Code Agents
di: Chen, Guhong, et al.
Pubblicazione: (2026)
di: Chen, Guhong, et al.
Pubblicazione: (2026)
VCM: Vision Concept Modeling Based on Implicit Contrastive Learning with Vision-Language Instruction Fine-Tuning
di: Luo, Run, et al.
Pubblicazione: (2025)
di: Luo, Run, et al.
Pubblicazione: (2025)
Marathon: A Race Through the Realm of Long Context with Large Language Models
di: Zhang, Lei, et al.
Pubblicazione: (2023)
di: Zhang, Lei, et al.
Pubblicazione: (2023)
OpenOmni: Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignment and Real-Time Self-Aware Emotional Speech Synthesis
di: Luo, Run, et al.
Pubblicazione: (2025)
di: Luo, Run, et al.
Pubblicazione: (2025)
Learning Ordinal Probabilistic Reward from Preferences
di: Chen, Longze, et al.
Pubblicazione: (2026)
di: Chen, Longze, et al.
Pubblicazione: (2026)
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
Small Language Model as Data Prospector for Large Language Model
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
di: Li, Siyi, et al.
Pubblicazione: (2026)
di: Li, Siyi, et al.
Pubblicazione: (2026)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
di: Luo, Run, et al.
Pubblicazione: (2025)
di: Luo, Run, et al.
Pubblicazione: (2025)
CollectiveSFT: Scaling Large Language Models for Chinese Medical Benchmark with Collective Instructions in Healthcare
di: Zhu, Jingwei, et al.
Pubblicazione: (2024)
di: Zhu, Jingwei, et al.
Pubblicazione: (2024)
RuCL: Stratified Rubric-Based Curriculum Learning for Multimodal Large Language Model Reasoning
di: Chen, Yukun, et al.
Pubblicazione: (2026)
di: Chen, Yukun, et al.
Pubblicazione: (2026)
EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection
di: Xu, Ancheng, et al.
Pubblicazione: (2025)
di: Xu, Ancheng, et al.
Pubblicazione: (2025)
How chromatin interactions shed light on interpreting non-coding genomic variants: opportunities and future direc-tions
di: Liang, Yuheng, et al.
Pubblicazione: (2024)
di: Liang, Yuheng, et al.
Pubblicazione: (2024)
Implicit Actor Critic Coupling via a Supervised Learning Framework for RLVR
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
IP-MOT: Instance Prompt Learning for Cross-Domain Multi-Object Tracking
di: Luo, Run, et al.
Pubblicazione: (2024)
di: Luo, Run, et al.
Pubblicazione: (2024)
MoZIP: A Multilingual Benchmark to Evaluate Large Language Models in Intellectual Property
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
Hierarchical Context Pruning: Optimizing Real-World Code Completion with Repository-Level Pretrained Code LLMs
di: Zhang, Lei, et al.
Pubblicazione: (2024)
di: Zhang, Lei, et al.
Pubblicazione: (2024)
Training Superior Sparse Autoencoders for Instruct Models
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
Structuring Reasoning for Complex Rules Beyond Flat Representations
di: Yang, Zhihao, et al.
Pubblicazione: (2025)
di: Yang, Zhihao, et al.
Pubblicazione: (2025)
Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective Ambiguity
di: Fang, Feiteng, et al.
Pubblicazione: (2025)
di: Fang, Feiteng, et al.
Pubblicazione: (2025)
Stereochemical Investigation of a Novel Biological Active Substance from the Secondary Metabolites of Marine Fungus Penicillium chrysogenum SYP-F-2720
di: Linwei Li
Pubblicazione: (2015)
di: Linwei Li
Pubblicazione: (2015)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
DEEM: Diffusion Models Serve as the Eyes of Large Language Models for Image Perception
di: Luo, Run, et al.
Pubblicazione: (2024)
di: Luo, Run, et al.
Pubblicazione: (2024)
Enhancing Monte Carlo Dropout Performance for Uncertainty Quantification
di: Asgharnezhad, Hamzeh, et al.
Pubblicazione: (2025)
di: Asgharnezhad, Hamzeh, et al.
Pubblicazione: (2025)
A High Conversion Gain and Low Phase Noise Self‐Oscillating Mixer Based on Positive Feedback Technique
di: Shiqiang Fu, et al.
Pubblicazione: (2025)
di: Shiqiang Fu, et al.
Pubblicazione: (2025)
Layer-wise Regularized Dropout for Neural Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FlowPIE: Test-Time Scientific Idea Evolution with Flow-Guided Literature Exploration
di: Wang, Qiyao, et al.
Pubblicazione: (2026) -
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
di: Chen, Guhong, et al.
Pubblicazione: (2024) -
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
di: Wang, Qiyao, et al.
Pubblicazione: (2024) -
PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination
di: Wang, Qiyao, et al.
Pubblicazione: (2026) -
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
di: Wang, Qiyao, et al.
Pubblicazione: (2026)