# R-LPS (FinancialPhraseBank) Benchmark Scores & Performance

> R-LPS (FinancialPhraseBank) has published results in 2 original benchmark tables. The evaluated configuration, source metrics, and precision remain visible below. These results do not produce a general model score or rank.

Canonical page: https://benchlm.ai/models/native-financialphrasebank-r-lps

Last updated: 2026-10-01

General benchmark catalog last updated: October 1, 2026. This profile’s source review has its own date above.

## Model Details

| Property | Value |
|----------|-------|
| Creator | Malo et al. |
| Source Type | Research system |
| Reasoning Type | Unspecified |
| Context Window | Not established by evaluation |
| Official model card | [Benchmark-owner results and configuration](https://arxiv.org/abs/1307.5336) |
| Overall Score | Not computed (source protocol results only) |
| Overall Rank | Unranked |

## Family & Coverage

- Family: R-LPS (FinancialPhraseBank)
- Variant: benchmark-system
- Benchmarks covered: 0 of 645
- Coverage note: Original benchmark result tables appear below; these metrics are separate from weighted benchmark slots.

## Original benchmark results

[All model results (JSON)](/api/data/benchmarks?model=native-financialphrasebank-r-lps) · [Numeric metrics (CSV)](/api/data/benchmarks?model=native-financialphrasebank-r-lps&format=csv)

### Class metrics at 100% and >75% agreement

The original study reports ten-fold cross-validation at four annotator-agreement thresholds. Its class-specific accuracy, precision, recall, and F1 are ratios from 0 to 1. These are not overall sentiment accuracy or Perplexity’s sampled panel. No unified modern leaderboard is published in the reviewed benchmark-owner sources.

All metrics are 0–1 ratios. SVM-MPQA is the paper’s baseline marked with footnote a.

Good Debt or Bad Debt: Detecting Semantic Orientations in Economic Texts — Pekka Malo and colleagues. Numeric results transcribed and reformatted; evaluation notes and model links added by BenchLM. Published numeric precision is retained.

[Published source](https://arxiv.org/abs/1307.5336) · [Full FinancialPhraseBank results](/benchmarks/financialphrasebank)

| Agreement | Class | Metric | R-LPS |
| --- | --- | --- | --- |
| 100% | Positive | Accuracy | 0.858 |
| 100% | Positive | Recall | 0.698 |
| 100% | Positive | Precision | 0.728 |
| 100% | Positive | F1-score | 0.713 |
| 100% | Neutral | Accuracy | 0.851 |
| 100% | Neutral | Recall | 0.887 |
| 100% | Neutral | Precision | 0.872 |
| 100% | Neutral | F1-score | 0.880 |
| 100% | Negative | Accuracy | 0.947 |
| 100% | Negative | Recall | 0.799 |
| 100% | Negative | Precision | 0.801 |
| 100% | Negative | F1-score | 0.800 |
| >75% | Positive | Accuracy | 0.826 |
| >75% | Positive | Recall | 0.614 |
| >75% | Positive | Precision | 0.677 |
| >75% | Positive | F1-score | 0.644 |
| >75% | Neutral | Accuracy | 0.799 |
| >75% | Neutral | Recall | 0.865 |
| >75% | Neutral | Precision | 0.821 |
| >75% | Neutral | F1-score | 0.842 |
| >75% | Negative | Accuracy | 0.939 |
| >75% | Negative | Recall | 0.707 |
| >75% | Negative | Precision | 0.771 |
| >75% | Negative | F1-score | 0.738 |

### Class metrics at >66% and >50% agreement

The original study reports ten-fold cross-validation at four annotator-agreement thresholds. Its class-specific accuracy, precision, recall, and F1 are ratios from 0 to 1. These are not overall sentiment accuracy or Perplexity’s sampled panel. No unified modern leaderboard is published in the reviewed benchmark-owner sources.

All metrics are 0–1 ratios. SVM-MPQA is the paper’s baseline marked with footnote a.

Good Debt or Bad Debt: Detecting Semantic Orientations in Economic Texts — Pekka Malo and colleagues. Numeric results transcribed and reformatted; evaluation notes and model links added by BenchLM. Published numeric precision is retained.

[Published source](https://arxiv.org/abs/1307.5336) · [Full FinancialPhraseBank results](/benchmarks/financialphrasebank)

| Agreement | Class | Metric | R-LPS |
| --- | --- | --- | --- |
| >66% | Positive | Accuracy | 0.799 |
| >66% | Positive | Recall | 0.592 |
| >66% | Positive | Precision | 0.651 |
| >66% | Positive | F1-score | 0.620 |
| >66% | Neutral | Accuracy | 0.761 |
| >66% | Neutral | Recall | 0.830 |
| >66% | Neutral | Precision | 0.786 |
| >66% | Neutral | F1-score | 0.807 |
| >66% | Negative | Accuracy | 0.931 |
| >66% | Negative | Recall | 0.681 |
| >66% | Negative | Precision | 0.731 |
| >66% | Negative | F1-score | 0.705 |
| >50% | Positive | Accuracy | 0.776 |
| >50% | Positive | Recall | 0.500 |
| >50% | Positive | Precision | 0.629 |
| >50% | Positive | F1-score | 0.557 |
| >50% | Neutral | Accuracy | 0.729 |
| >50% | Neutral | Recall | 0.836 |
| >50% | Neutral | Precision | 0.741 |
| >50% | Neutral | F1-score | 0.786 |
| >50% | Negative | Accuracy | 0.922 |
| >50% | Negative | Recall | 0.613 |
| >50% | Negative | Precision | 0.721 |
| >50% | Negative | F1-score | 0.662 |

## Other Malo et al. Models

- [LPS (FinancialPhraseBank)](/models/native-financialphrasebank-lps) - Score: not computed
- [SVM-MPQA (FinancialPhraseBank)](/models/native-financialphrasebank-svm-mpqa) - Score: not computed
- [W-Loughran (FinancialPhraseBank)](/models/native-financialphrasebank-w-loughran) - Score: not computed
- [W-MPQA (FinancialPhraseBank)](/models/native-financialphrasebank-w-mpqa) - Score: not computed
