paper-with-me

홈 › Papers

MarianCG: a code generation transformer model inspired by machine translation

2022-11-22 · Journal of Engineering and Applied Science 2022 11 · Ahmed S. Soliman, Mayada M. Hadhoud, Samir I. Shaheen

The idea that computers can build their own programs is extremely significant, and many researchers are working on this challenge. Code generation is described as the process of generating executable code that can be run directly on the computer and fulfills the natural language requirements. It is an intriguing topic that might assist developers to learn a new software technology or programming language, or it could be a simple technique to help in coding through the description of the natural language code developer. In this paper, we present MarianCG, a code generation Transformer model used to tackle the code generation challenge of generating python code from natural language descriptions. Marian neural machine translation (NMT), which is the core model of the Microsoft Translator, is the basis for our NL-to-Code translation engine and is the heart of the teaching model. MarianMT is the teacher language model in our study, and it is one of the most successful machine translation transformers. In our approach, we use a sinusoidal positional embedding technique to represent the position of each token in the text, as well as no layer normalization embedding. Our code generation approach, MarianCG, is based on fine-tuning a machine translation pre-trained language model. This allows us to demonstrate that the pre-trained translation model can also operate and work as a code generation model. The proposed model outperforms recent state-of-the-art models in the problem of code generation when trained on the CoNaLa and DJANGO datasets. MarianCG model scores a BLEU score of 34.43 and an exact match accuracy of 10.2% on the CoNaLa dataset. Also, this model records a BLEU score of 90.41 and an exact match accuracy of 81.83% on the DJANGO dataset. The implementation of MarianCG model and relevant resources are available at https://www.github.com/AhmedSSoliman/MarianCG-NL-to-Code.

📄 PDF Abstract BibTeX

Code (1)

AhmedSSoliman/MarianCG-NL-to-Code 공식 구현

Tasks

Code GenerationCode TranslationLanguage ModelingLanguage ModellingMachine TranslationNMTTranslation

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Residual Connection 설명 없음
Adam 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…

Similar Papers 제목 키워드 기반

Surgical Instruction Generation with Transformers

2021-07-14 · Jinglu Zhang, Yinyu Nie, Jian Chang, Jian Jun Zhang

Automatic surgical instruction generation is a prerequisite towards intra-operative context-aware surgical assistance. However, generating instructions from surgical scenes is challenging, as it requires jointly understa…

DecoderMachine TranslationReinforcement Learning (RL)Translation

ODE Transformer: An Ordinary Differential Equation-Inspired Model for Sequence Generation

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Residual networks are an Euler discretization of solutions to Ordinary Differential Equations (ODE). This paper explores a deeper relationship between Transformer and numerical ODE methods. We first show that a residual …

Abstractive Text SummarizationMachine TranslationTranslation

ODE Transformer: An Ordinary Differential Equation-Inspired Model for Sequence Generation

2022-03-17 · ACL 2022 5 · Bei Li, Quan Du, Tao Zhou, Yi Jing 외

Residual networks are an Euler discretization of solutions to Ordinary Differential Equations (ODE). This paper explores a deeper relationship between Transformer and numerical ODE methods. We first show that a residual …

Abstractive Text SummarizationMachine TranslationTranslation

TransTIC: Transferring Transformer-based Image Compression from Human Perception to Machine Perception

2023-06-08 · ICCV 2023 1 · Yi-Hsin Chen, Ying-Chieh Weng, Chia-Hao Kao, Cheng Chien 외

This work aims for transferring a Transformer-based image compression codec from human perception to machine perception without fine-tuning the codec. We propose a transferable Transformer-based image compression framewo…

DecoderImage CompressionVisual Prompt Tuning

POSESTITCH-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation

2025-10-31 · Abhinav Joshi, Vaibhav Sharma, Sanjeet Singh, Ashutosh Modi arxiv

Sign language translation remains a challenging task due to the scarcity of large-scale, sentence-aligned datasets. Prior arts have focused on various feature extraction and architectural changes to support neural machin…

Sign Language TranslationMachine Translation