Methods › Natural Language Processing › Language Models › OPT › Papers, page 3
OPT
Papers archive 2025-07-28
archive papers tagged: 285 · with a code link: 123 · where Syntology ran a sample: 55 (41 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 285 tagged: 41 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument)
Page 3 of 3: papers 201 to 285 of 285, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Active Learning Principles for In-Context Learning with Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14264
-
Goat: Fine-tuned LLaMA Outperforms GPT-4 on Arithmetic Tasks 23 May 2023 · 1 repository · arXiv:2305.14201
-
Can ChatGPT Detect Intent? Evaluating Large Language Models for Spoken Language Understanding 22 May 2023 · 0 repositories · arXiv:2305.13512
-
OPT-R: Exploring the Role of Explanations in Finetuning and Prompting for Reasoning Skills of Large Language Models 19 May 2023 · 0 repositories · arXiv:2305.12001
-
Scaling laws for language encoding models in fMRI 19 May 2023 · 1 repository · arXiv:2305.11863
-
Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility 15 May 2023 · 1 repository · arXiv:2305.10235
-
Knowledge Rumination for Pre-trained Language Models 15 May 2023 · 1 repository · arXiv:2305.08732
-
Towards Versatile and Efficient Visual Knowledge Integration into Pre-trained Language Models with Cross-Modal Adapters 12 May 2023 · 0 repositories · arXiv:2305.07358
-
"When Words Fail, Emojis Prevail": Generating Sarcastic Utterances with Emoji Using Valence Reversal and Semantic Incongruity 6 May 2023 · 0 repositories · arXiv:2305.04105
-
Conformal Nucleus Sampling 4 May 2023 · 0 repositories · arXiv:2305.02633
-
TextDeformer: Geometry Manipulation using Text Guidance 26 Apr 2023 · 1 repository · arXiv:2304.13348Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
Better Question-Answering Models on a Budget 24 Apr 2023 · 1 repository · arXiv:2304.12370
-
How 'one-size-fits-all' public works contract does it better? An assessment of infrastructure provision in Italy 21 Apr 2023 · 0 repositories · arXiv:2304.10776
-
Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling 18 Apr 2023 · 1 repository · arXiv:2304.09145Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
LongForm: Effective Instruction Tuning with Reverse Instructions 17 Apr 2023 · 2 repositories · arXiv:2304.08460
-
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis 10 Apr 2023 · 2 repositories · arXiv:2304.04675
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Scalable and Accurate Self-supervised Multimodal Representation Learning without Aligned Video and Text Data 4 Apr 2023 · 0 repositories · arXiv:2304.02080
-
Koala: An Index for Quantifying Overlaps with Pre-training Corpora 26 Mar 2023 · 0 repositories · arXiv:2303.14770Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Freestyle Layout-to-Image Synthesis 25 Mar 2023 · 1 repository · arXiv:2303.14412Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Uncertainty Aware Active Learning for Reconfiguration of Pre-trained Deep Object-Detection Networks for New Target Domains 22 Mar 2023 · 0 repositories · arXiv:2303.12760
-
DC-CCL: Device-Cloud Collaborative Controlled Learning for Large Vision Models 18 Mar 2023 · 0 repositories · arXiv:2303.10361
-
A Dynamic Multi-Scale Voxel Flow Network for Video Prediction 17 Mar 2023 · 1 repository · arXiv:2303.09875Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Fusing Visual Appearance and Geometry for Multi-modality 6DoF Object Tracking 22 Feb 2023 · 1 repository · arXiv:2302.11458
-
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints 17 Feb 2023 · 1 repository · arXiv:2302.09185
-
Learning with Rejection for Abstractive Text Summarization 16 Feb 2023 · 1 repository · arXiv:2302.08531
-
OPT: One-shot Pose-Controllable Talking Head Generation 16 Feb 2023 · 0 repositories · arXiv:2302.08197
-
Making Substitute Models More Bayesian Can Enhance Transferability of Adversarial Examples 10 Feb 2023 · 1 repository · arXiv:2302.05086Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Participatory Personalization in Classification 8 Feb 2023 · 0 repositories · arXiv:2302.03874
-
Linear Optimal Partial Transport Embedding 7 Feb 2023 · 1 repository · arXiv:2302.03232
-
Towards a User Privacy-Aware Mobile Gaming App Installation Prediction Model 7 Feb 2023 · 0 repositories · arXiv:2302.03332
-
YOLO-based Object Detection in Industry 4.0 Fischertechnik Model Environment 30 Jan 2023 · 0 repositories · arXiv:2301.12827
-
Understanding the Effectiveness of Very Large Language Models on Dialog Evaluation 27 Jan 2023 · 0 repositories · arXiv:2301.12004
-
Optimus-CC: Efficient Large NLP Model Training with 3D Parallelism Aware Communication Compression 24 Jan 2023 · 0 repositories · arXiv:2301.09830
-
Automatic Standardization of Arabic Dialects for Machine Translation 9 Jan 2023 · 0 repositories · arXiv:2301.03447
-
Vector Quantization With Self-Attention for Quality-Independent Representation Learning 1 Jan 2023 · 0 repositories
-
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times? 23 Dec 2022 · 0 repositories · arXiv:2212.12131
-
OPT-IML: Scaling Language Model Instruction Meta Learning through the Lens of Generalization 22 Dec 2022 · 1 repository · arXiv:2212.12017
-
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10461
-
The case for 4-bit precision: k-bit Inference Scaling Laws 19 Dec 2022 · 1 repository · arXiv:2212.09720
-
Training Trajectories of Language Models Across Scales 19 Dec 2022 · 1 repository · arXiv:2212.09803
-
Sliced Optimal Partial Transport 15 Dec 2022 · 2 repositories · arXiv:2212.08049
-
Elixir: Train a Large Language Model on a Small GPU Cluster 10 Dec 2022 · 2 repositories · arXiv:2212.05339
-
Distributed Stochastic Gradient Descent with Cost-Sensitive and Strategic Agents 5 Dec 2022 · 0 repositories · arXiv:2212.02049
-
On the Design of Communication-Efficient Federated Learning for Health Monitoring 30 Nov 2022 · 0 repositories · arXiv:2211.16952
-
Beyond S-curves: Recurrent Neural Networks for Technology Forecasting 28 Nov 2022 · 0 repositories · arXiv:2211.15334
-
SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models 18 Nov 2022 · 5 repositories · arXiv:2211.10438Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
Evaluating the Factual Consistency of Large Language Models Through News Summarization 15 Nov 2022 · 1 repository · arXiv:2211.08412
-
On the Compositional Generalization Gap of In-Context Learning 15 Nov 2022 · 0 repositories · arXiv:2211.08473
-
PriMask: Cascadable and Collusion-Resilient Data Masking for Mobile Cloud Inference 12 Nov 2022 · 1 repository · arXiv:2211.06716
-
Towards real-time 6D pose estimation of objects in single-view cone-beam X-ray 6 Nov 2022 · 0 repositories · arXiv:2211.03211
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
On the Transformation of Latent Space in Fine-Tuned NLP Models 23 Oct 2022 · 0 repositories · arXiv:2210.12696
-
SMaLL-100: Introducing Shallow Multilingual Machine Translation Model for Low-Resource Languages 20 Oct 2022 · 3 repositories · arXiv:2210.11621Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Beyond capacity: contractual form in electricity reliability obligations 19 Oct 2022 · 0 repositories · arXiv:2210.10858
-
lo-fi: distributed fine-tuning without communication 19 Oct 2022 · 0 repositories · arXiv:2210.11948
-
MaSS: Multi-attribute Selective Suppression 18 Oct 2022 · 0 repositories · arXiv:2210.09904
-
That's the Wrong Lung! Evaluating and Improving the Interpretability of Unsupervised Multimodal Encoders for Medical Data 12 Oct 2022 · 0 repositories · arXiv:2210.06565
-
Autonomous Asteroid Characterization Through Nanosatellite Swarming 11 Oct 2022 · 0 repositories · arXiv:2210.05518
-
AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models 8 Oct 2022 · 0 repositories · arXiv:2210.03858
-
SpaceQA: Answering Questions about the Design of Space Missions and Space Craft Concepts 7 Oct 2022 · 1 repository · arXiv:2210.03422
-
Effective Self-supervised Pre-training on Low-compute Networks without Distillation 6 Oct 2022 · 1 repository · arXiv:2210.02808Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Ask Me Anything: A simple strategy for prompting language models 5 Oct 2022 · 3 repositories · arXiv:2210.02441Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Language Models are Pretty Good Zero-Shot Video Game Bug Detectors 5 Oct 2022 · 1 repository · arXiv:2210.02506
-
Recitation-Augmented Language Models 4 Oct 2022 · 1 repository · arXiv:2210.01296Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Data-driven Approach to Named Entity Recognition for Early Modern French 1 Oct 2022 · 0 repositories
-
Optimal Partial Transport Based Sentence Selection for Long-form Document Matching 1 Oct 2022 · 1 repository
-
Pre-trained Speech Representations as Feature Extractors for Speech Quality Assessment in Online Conferencing Applications 1 Oct 2022 · 1 repository · arXiv:2210.00259
-
Moral Mimicry: Large Language Models Produce Moral Rationalizations Tailored to Political Identity 24 Sep 2022 · 0 repositories · arXiv:2209.12106
-
An Improved Algorithm For Online Min-Sum Set Cover 11 Sep 2022 · 0 repositories · arXiv:2209.04870
-
FOLIO: Natural Language Reasoning with First-Order Logic 2 Sep 2022 · 1 repository · arXiv:2209.00840
-
Effective Class-Imbalance learning based on SMOTE and Convolutional Neural Networks 1 Sep 2022 · 0 repositories · arXiv:2209.00653
-
A New Kind of Adversarial Example 4 Aug 2022 · 1 repository · arXiv:2208.02430
-
RenderNet: Visual Relocalization Using Virtual Viewpoints in Large-Scale Indoor Environments 26 Jul 2022 · 0 repositories · arXiv:2207.12579
-
Smooth Anonymity for Sparse Graphs 13 Jul 2022 · 0 repositories · arXiv:2207.06358
-
Interaction Pattern Disentangling for Multi-Agent Reinforcement Learning 8 Jul 2022 · 1 repository · arXiv:2207.03902
-
Scalable Differentially Private Clustering via Hierarchically Separated Trees 17 Jun 2022 · 1 repository · arXiv:2206.08646
-
From Human Days to Machine Seconds: Automatically Answering and Generating Machine Learning Final Exams 11 Jun 2022 · 0 repositories · arXiv:2206.05442
-
Contextual Bandits with Knapsacks for a Conversion Model 1 Jun 2022 · 0 repositories · arXiv:2206.00314
-
Controlling Extra-Textual Attributes about Dialogue Participants: A Case Study of English-to-Polish Neural Machine Translation 1 Jun 2022 · 0 repositories
-
Variation in the Expression and Annotation of Emotions: A Wizard of Oz Pilot Study 1 Jun 2022 · 0 repositories
-
Controlling Extra-Textual Attributes about Dialogue Participants -- A Case Study of English-to-Polish Neural Machine Translation 10 May 2022 · 0 repositories · arXiv:2205.04747
-
The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning 6 May 2022 · 1 repository · arXiv:2205.03401
-
Weakly Supervised 3D Point Cloud Segmentation via Multi-Prototype Learning 6 May 2022 · 0 repositories · arXiv:2205.03137
-
OPT: Open Pre-trained Transformer Language Models 2 May 2022 · 11 repositories · arXiv:2205.01068Syntology official (archive's flag): 9 ran · 14 ran (of which 2 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 17 pointer-only (licence)