Text-to-Code Generation with Modality-relative Pre-training
Fuente:
arXiv
Saved in:
| Main Authors: | Christopoulou, Fenia, Zhang, Guchun, Lampouras, Gerasimos |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
by: Christopoulou, Fenia, et al.
Published: (2024)
by: Christopoulou, Fenia, et al.
Published: (2024)
Human-inspired Episodic Memory for Infinite Context LLMs
by: Fountas, Zafeirios, et al.
Published: (2024)
by: Fountas, Zafeirios, et al.
Published: (2024)
Code-Optimise: Self-Generated Preference Data for Correctness and Efficiency
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition
by: Li, Yupei, et al.
Published: (2025)
by: Li, Yupei, et al.
Published: (2025)
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
by: Gritta, Milan, et al.
Published: (2024)
by: Gritta, Milan, et al.
Published: (2024)
DReSD: Dense Retrieval for Speculative Decoding
by: Gritta, Milan, et al.
Published: (2025)
by: Gritta, Milan, et al.
Published: (2025)
DRIFT: Decompose, Retrieve, Illustrate, then Formalize Theorems
by: Zhang, Meiru, et al.
Published: (2025)
by: Zhang, Meiru, et al.
Published: (2025)
Conjecturing: An Overlooked Step in Formal Mathematical Reasoning
by: Sivakumar, Jasivan Alex, et al.
Published: (2025)
by: Sivakumar, Jasivan Alex, et al.
Published: (2025)
Findings of the First Workshop on Simulating Conversational Intelligence in Chat
by: Graham, Yvette, et al.
Published: (2024)
by: Graham, Yvette, et al.
Published: (2024)
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024)
by: Zimmer, Matthieu, et al.
Published: (2024)
To Code, or Not To Code? Exploring Impact of Code in Pre-training
by: Aryabumi, Viraat, et al.
Published: (2024)
by: Aryabumi, Viraat, et al.
Published: (2024)
Cross-Modal Robustness Transfer (CMRT): Training Robust Speech Translation Models Using Adversarial Text
by: Issam, Abderrahmane, et al.
Published: (2026)
by: Issam, Abderrahmane, et al.
Published: (2026)
Encoder-Decoder Framework for Interactive Free Verses with Generation with Controllable High-Quality Rhyming
by: Pasini, Tommaso, et al.
Published: (2024)
by: Pasini, Tommaso, et al.
Published: (2024)
RecGPT: Generative Pre-training for Text-based Recommendation
by: Ngo, Hoang, et al.
Published: (2024)
by: Ngo, Hoang, et al.
Published: (2024)
Visually Guided Generative Text-Layout Pre-training for Document Intelligence
by: Mao, Zhiming, et al.
Published: (2024)
by: Mao, Zhiming, et al.
Published: (2024)
Cued Speech Generation Leveraging a Pre-trained Audiovisual Text-to-Speech Model
by: Sankar, Sanjana, et al.
Published: (2025)
by: Sankar, Sanjana, et al.
Published: (2025)
Order-Based Pre-training Strategies for Procedural Text Understanding
by: Nandy, Abhilash, et al.
Published: (2024)
by: Nandy, Abhilash, et al.
Published: (2024)
Automated Detection of Pre-training Text in Black-box LLMs
by: Hu, Ruihan, et al.
Published: (2025)
by: Hu, Ruihan, et al.
Published: (2025)
Enhancing EEG-to-Text Decoding through Transferable Representations from Pre-trained Contrastive EEG-Text Masked Autoencoder
by: Wang, Jiaqi, et al.
Published: (2024)
by: Wang, Jiaqi, et al.
Published: (2024)
Towards Versatile and Efficient Visual Knowledge Integration into Pre-trained Language Models with Cross-Modal Adapters
by: Zhang, Xinyun, et al.
Published: (2023)
by: Zhang, Xinyun, et al.
Published: (2023)
A Survey of Pre-trained Language Models for Processing Scientific Text
by: Ho, Xanh, et al.
Published: (2024)
by: Ho, Xanh, et al.
Published: (2024)
Text Embeddings by Weakly-Supervised Contrastive Pre-training
by: Wang, Liang, et al.
Published: (2022)
by: Wang, Liang, et al.
Published: (2022)
PhoGPT: Generative Pre-training for Vietnamese
by: Nguyen, Dat Quoc, et al.
Published: (2023)
by: Nguyen, Dat Quoc, et al.
Published: (2023)
DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment
by: Issam, Abderrahmane, et al.
Published: (2025)
by: Issam, Abderrahmane, et al.
Published: (2025)
Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation
by: Zhang, Yuhui, et al.
Published: (2023)
by: Zhang, Yuhui, et al.
Published: (2023)
Evaluating LLMs and Pre-trained Models for Text Summarization Across Diverse Datasets
by: Rehman, Tohida, et al.
Published: (2025)
by: Rehman, Tohida, et al.
Published: (2025)
CodeSCM: Causal Analysis for Multi-Modal Code Generation
by: Gupta, Mukur, et al.
Published: (2025)
by: Gupta, Mukur, et al.
Published: (2025)
Scaling Speech-Text Pre-training with Synthetic Interleaved Data
by: Zeng, Aohan, et al.
Published: (2024)
by: Zeng, Aohan, et al.
Published: (2024)
Code-Mixed Probes Show How Pre-Trained Models Generalise On Code-Switched Text
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
Unveiling the Deficiencies of Pre-trained Text-and-Layout Models in Real-world Visually-rich Document Information Extraction
by: Zhang, Chong, et al.
Published: (2024)
by: Zhang, Chong, et al.
Published: (2024)
Fine-tuning Pre-trained Language Models for Few-shot Intent Detection: Supervised Pre-training and Isotropization
by: Zhang, Haode, et al.
Published: (2022)
by: Zhang, Haode, et al.
Published: (2022)
Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text
by: Huang, Chengyu, et al.
Published: (2026)
by: Huang, Chengyu, et al.
Published: (2026)
Pre-trained Language Models Return Distinguishable Probability Distributions to Unfaithfully Hallucinated Texts
by: Cha, Taehun, et al.
Published: (2024)
by: Cha, Taehun, et al.
Published: (2024)
The Efficiency of Pre-training with Objective Masking in Pseudo Labeling for Semi-Supervised Text Classification
by: Hatefi, Arezoo, et al.
Published: (2025)
by: Hatefi, Arezoo, et al.
Published: (2025)
MEDVOC: Vocabulary Adaptation for Fine-tuning Pre-trained Language Models on Medical Text Summarization
by: Balde, Gunjan, et al.
Published: (2024)
by: Balde, Gunjan, et al.
Published: (2024)
TiMix: Text-aware Image Mixing for Effective Vision-Language Pre-training
by: Jiang, Chaoya, et al.
Published: (2023)
by: Jiang, Chaoya, et al.
Published: (2023)
A Benchmark for Deep Information Synthesis
by: Paul, Debjit, et al.
Published: (2026)
by: Paul, Debjit, et al.
Published: (2026)
Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch
by: Strazda, Elza, et al.
Published: (2025)
by: Strazda, Elza, et al.
Published: (2025)
Structure-aware Fine-tuning for Code Pre-trained Models
by: Wu, Jiayi, et al.
Published: (2024)
by: Wu, Jiayi, et al.
Published: (2024)
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training
by: Kim, Sanghwan, et al.
Published: (2024)
by: Kim, Sanghwan, et al.
Published: (2024)
Similar Items
-
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
by: Christopoulou, Fenia, et al.
Published: (2024) -
Human-inspired Episodic Memory for Infinite Context LLMs
by: Fountas, Zafeirios, et al.
Published: (2024) -
Code-Optimise: Self-Generated Preference Data for Correctness and Efficiency
by: Gee, Leonidas, et al.
Published: (2024) -
TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition
by: Li, Yupei, et al.
Published: (2025) -
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
by: Gritta, Milan, et al.
Published: (2024)