PERSLIS SCIENCE / INSTRUMENTS

One question in.
Your scientific stack answers together.

Connect literature, molecular databases, computation, files and lab systems behind one source contract. Keep the tools your lab trusts; lose the copy-paste research trail.

◈ Evidence connected⌁ Unknowns stay visible
YOUR STACK, CONNECTED

Keep the instruments. Replace the fragmentation.

PubMedPMCUniProtRCSB PDBAlphaFoldEnsemblClinVarPubChemChEMBLKEGGReactomeInterPro
◈SOURCE CONTRACT
One research recordGraph + ledger + notebook + runs
For scientists

Ask one question across the stack and inspect the exact records behind the answer.

For labs

Connect existing services and scoped instruments without a rip-and-replace migration.

THE COMPLETE SCIENCE TOOLSET

Every instrument. One connected page.

Browse every public scientific service, the exact console tools it powers, how Perslis connects it, and a real source-pinned result. Expand any instrument without leaving this page.

44
instruments
19
services
7
benches
01
INSTRUMENT BENCH

Literature

2 services
LITERATUREPubMedBiomedical literature search — abstracts pinned to PMID, DOI and citation.pubmed_search · pubmed_abstracts
THE INTEGRATION

How Perslis integrates PubMed

PubMed is the National Library of Medicine's index of biomedical literature — abstracts, MeSH terms, DOIs and citation metadata across MEDLINE. It is the first place most research questions land.

On the console, pubmed_search runs a relevance- or date-sorted query and returns structured records; pubmed_abstracts fetches the full abstract text by PMID. Both hand back the identifiers a claim can be traced by.

The failure mode this fixes is the fabricated citation — the plausible paper that does not exist. Because Lois can only report what PubMed returned, every reference on the console carries a live PMID that resolves. There is no path for a made-up DOI to enter the record.

PubMed is also the hub of the literature mesh: a PMID that carries a PMCID hands straight to pmc_fulltext for open-access full text, and the same question fans out to OpenAlex and Semantic Scholar so a search is a triangulation, not a single opinion.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · pubmed_search("sickle cell disease hemoglobin therapy")
Top resultSickle Cell Disease.
AuthorsPecker LH, Lanzkron S
JournalAnn Intern Med, 2021
PMID33428443
DOI10.7326/AITC202101190
Total matches4,519

pubmed · PMID 33428443 · retrieved 2026-09-18

01

For AI agents

A PubMed MCP-style tool an LLM can call directly — structured records in, no scraping, no key juggling.

02

One question, every source

The same query hits PubMed, OpenAlex and Semantic Scholar at once, then cross-references PMCID to full text.

03

No hallucinated citations

Every reference resolves to a real PMID. A claim that cannot name its source is refused by construction.

LITERATUREPubMed CentralOpen-access full text, fetched by PMCID.pmc_fulltext
THE INTEGRATION

How Perslis integrates PubMed Central

PubMed Central (PMC) is the NIH free full-text archive of biomedical literature. Where a paper is open access, PMC has the complete article — introduction, methods, figures legends, results and references.

On the console, pmc_fulltext takes a PMCID and returns the article body. It is the second half of a PubMed hit: search finds the paper, PMC reads it.

The handoff is automatic. When pubmed_search returns a record that carries a PMCID, that identifier is the key straight into full text — no second search, no guessing a URL. The claim and the paragraph it came from stay one click apart.

This is what makes grounded reading possible offline too: full text pulled through PMC can be harvested into a local topic pack, so a review still has the source article when the cable is pulled.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive handoff · a PubMed hit carrying a PMCID
ArticleHematopoietic Stem Cell Gene-Addition/Editing Therapy in Sickle Cell Disease
JournalCells, 2022
PMID35681538
PMCIDPMC9180595
Full textfetched by pmc_fulltext("PMC9180595")

pmc · PMC9180595 · retrieved 2026-09-18

01

Abstract to article, in one step

A PMCID from a PubMed hit resolves straight to full text — the console closes the loop for you.

02

Grounded reading for agents

An LLM reasons over the real methods and results, not a truncated abstract or a paraphrase.

03

Survives the cable being pulled

Open-access full text can be harvested into a local pack so the source paper stays with the record offline.

02
INSTRUMENT BENCH

Scholarly Graph

4 services
SCHOLARLY GRAPHarXivPreprints across quantitative biology and beyond.arxiv_search
THE INTEGRATION

How Perslis integrates arXiv

arXiv is the open preprint server for physics, mathematics, computer science and quantitative biology (q-bio). A paper is here before, and often instead of, peer review.

On the console, arxiv_search runs a query, optionally scoped to a category such as q-bio.GN, and returns titles, authors, abstracts and stable arXiv identifiers that link straight to the PDF.

Preprints are where the risk of hallucination is highest — the work is new, so a model is most tempted to invent a plausible-sounding one. Routing through arxiv_search removes the temptation: every preprint the console cites has an arXiv ID that resolves to a real abstract and PDF.

arXiv sits alongside bioRxiv on the preprint bench and beside OpenAlex and Semantic Scholar on the scholarly graph, so a frontier question is answered from preprints and the published record at once.

Open the source ↗
REAL, SOURCE-PINNED RESULTarxiv_search · scope q-bio, pinned to arXiv ID + PDF
Querydeep learning in genomics
Scopecategory q-bio.GN
Returnstitle, authors, abstract, arXiv ID
Each hitresolves to /abs/<id> and the PDF
Contractno arXiv ID, no citation

arxiv · q-bio.GN · queried via the export API under the source contract

01

The frontier, callable

An LLM reaches preprints directly — the same interface as PubMed, no separate scraper.

02

Preprint plus published

arXiv and bioRxiv answer together with the peer-reviewed record from OpenAlex and Semantic Scholar.

03

Every preprint resolves

A cited preprint always carries an arXiv ID that opens the real abstract and PDF.

SCHOLARLY GRAPHbioRxiv · medRxivPreprint feeds by date window; DOI fetch with the published version when one exists.biorxiv_recent · biorxiv_fetch
THE INTEGRATION

How Perslis integrates bioRxiv & medRxiv

bioRxiv (biology) and medRxiv (clinical/health) are the life-science preprint servers. Their API is date-window based — there is no keyword search — so you browse what posted, newest first, and filter by category.

On the console, biorxiv_recent returns a window's preprints (optionally filtered to a category like genomics), and biorxiv_fetch retrieves a specific DOI, following through to the published article when one exists.

The runtime is honest about the API's shape. bioRxiv has no keyword search, so the instrument does not pretend to: it returns exactly the date window asked for, filters category client-side, and states plainly when a window is empty rather than inventing preprints to fill it.

When a preprint has been published, biorxiv_fetch follows the DOI to the version of record — so a citation points at the peer-reviewed paper, not a superseded draft, the moment one exists.

Open the source ↗
REAL, SOURCE-PINNED RESULTbiorxiv_recent · date-window feed, honest about empty windows
Callbiorxiv_recent(days=3, category=genomics)
Window2026-09-15 .. 2026-09-18
API shapedate-window only — no keyword search
Empty windowreturned verbatim as 0, not padded
biorxiv_fetchDOI → published version when it exists

biorxiv · details API · retrieved 2026-09-18

01

Preprint feed, callable

An agent watches the newest biology by date window without scraping the site.

02

Preprint to version of record

A DOI follows through to the published paper the moment the preprint is accepted.

03

Honest about empty

An empty window returns as zero — the runtime never fabricates preprints to fill a gap.

SCHOLARLY GRAPHOpenAlexOpen scholarly metadata — works, citation counts, any discipline.openalex_search · openalex_work
THE INTEGRATION

How Perslis integrates OpenAlex

OpenAlex is a free, open replacement for the old proprietary citation databases. It covers 250 million-plus works with authorship, venue, year and citation metadata — biomedicine and everything beyond it.

On the console, openalex_search queries the corpus and openalex_work resolves a single work by ID, returning the metadata that lets an agent weigh how established an idea is.

OpenAlex is the runtime's discipline-agnostic index: PubMed covers biomedicine, but a materials or physics question needs a broader map, and OpenAlex provides it without a paywall or a key. Every work carries an OpenAlex ID and, where known, a DOI — both resolve.

It answers next to Semantic Scholar, whose forward-citation graph complements OpenAlex's breadth, so a single question yields both how much a paper is cited and who built on it.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · openalex_search("highly accurate protein structure prediction with AlphaFold")
Top workHighly accurate protein structure prediction with AlphaFold
Lead authorsJohn Jumper, Demis Hassabis, et al.
VenueNature, 2021
Cited by47,403
OpenAlex IDW3177828909
DOI10.1038/s41586-021-03819-2

openalex · W3177828909 · retrieved 2026-09-18

01

Every discipline, no key

One open index of all scholarship — an agent maps a field without a subscription.

02

Breadth plus citation depth

OpenAlex's coverage pairs with Semantic Scholar's forward-citation graph in one answer.

03

Every work resolves

OpenAlex ID and DOI both open the real record — citation counts you can check.

SCHOLARLY GRAPHSemantic ScholarPapers plus the forward citation graph — who built on what.semantic_scholar_search · semantic_scholar_citations
THE INTEGRATION

How Perslis integrates Semantic Scholar

Semantic Scholar is the Allen Institute's scholarly graph — papers, abstracts, citation counts and, distinctively, the forward citation edges that let you trace a line of work forward in time.

On the console, semantic_scholar_search queries papers and semantic_scholar_citations returns the works that cite a given paper — the shortest path to what happened next.

Its unique value on the console is direction. Given a landmark paper — say AlphaFold's Nature article, DOI 10.1038/s41586-021-03819-2 — the citation instrument walks forward to the work that built on it, turning a single reference into a live research front an agent can follow.

The unauthenticated tier is rate-limited, and the runtime treats that honestly: a throttled call is reported as unverified, never silently dropped or filled with a guessed number. It pairs with OpenAlex, whose open breadth backs up any gap.

Open the source ↗
REAL, SOURCE-PINNED RESULTsemantic_scholar_citations · forward graph, anchored on a real DOI
Anchor paperHighly accurate protein structure prediction with AlphaFold
VenueNature, 2021
DOI10.1038/s41586-021-03819-2
Instrumentsemantic_scholar_citations → who cited it
Rate-limited callreported as unverified, never faked

semantic scholar · via DOI · unauthenticated tier is rate-limited

01

Citation graph, callable

An agent walks forward from a paper to what built on it — a research front, not a static reference list.

02

Depth beside breadth

Semantic Scholar's forward edges complement OpenAlex's discipline-wide coverage in one answer.

03

Honest under rate limits

A throttled response is marked unverified — the runtime never invents a citation count to look complete.

03
INSTRUMENT BENCH

Proteins & Structures

4 services
PROTEINS & STRUCTURESUniProtCurated proteins — function, disease links, domains, sequence.uniprot_search · uniprot_entry
THE INTEGRATION

How Perslis integrates UniProt

UniProt (UniProtKB) is the expert-curated protein knowledgebase — reviewed entries with function, subcellular location, disease associations, sequence features and cross-references to structure and genome databases.

On the console, uniprot_search resolves a name or gene to accessions, and uniprot_entry returns the full record — function, disease, domains and its list of PDB structures — pinned to the accession.

UniProt is the join key of the protein bench. Ask about hemoglobin and Lois resolves it to P69905, then that one accession fans out: its PDB cross-references become structure lookups, its sequence becomes an AlphaFold model, its features become InterPro domains. The instruments talk to each other because they share the UniProt anchor.

Curated disease text — alpha-thalassemia, Heinz body anemia — is reported verbatim from the entry, never paraphrased into something a model finds tidier. If UniProt says it, the console says it; if it does not, the console does not invent it.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · uniprot_entry("P69905")
EntryHBA_HUMAN — Hemoglobin subunit alpha
OrganismHomo sapiens (reviewed)
GenesHBA1, HBA2
Length142 aa
DomainGlobin (residues 2–142)
DiseaseAlpha-thalassemia; Heinz body anemia
PDB structures300+ cross-referenced (e.g. 1HHO)

uniprot · P69905 · retrieved 2026-09-18

01

The join key for proteins

One accession resolves against PDB, AlphaFold, InterPro and Ensembl — no ID copied between tabs.

02

Curated, callable

An LLM reaches reviewed function and disease text directly, structured, pinned to accession.

03

Verbatim, not paraphrased

Disease and function text is reported as UniProt wrote it — the console never tidies it into a fabrication.

PROTEINS & STRUCTURESRCSB PDBExperimental structures — method, resolution, primary citation.pdb_search · pdb_entry
THE INTEGRATION

How Perslis integrates RCSB PDB

RCSB PDB is the worldwide repository of macromolecular structures solved by X-ray crystallography, cryo-EM and NMR. Each entry records how it was determined and at what resolution.

On the console, pdb_search finds structures and pdb_entry returns a single one — method, resolution, molecular weight, primary citation (with DOI and PMID) and direct .pdb/.cif download links.

PDB is where a protein's structure claim becomes checkable. A UniProt entry lists its cross-referenced PDB IDs; the console follows one straight to pdb_entry, and the resolution and primary citation come back with it — so a structural statement always carries the evidence and the paper behind it.

Experimental structure sits deliberately beside AlphaFold's prediction on the same bench: an agent can compare a measured structure against a predicted one and see, from the metadata, which is which.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · pdb_entry("1HHO")
TitleStructure of human oxyhaemoglobin at 2.1 Å resolution
MethodX-ray diffraction
Resolution2.1 Å
Deposited1983-06-10
Primary citationJ Mol Biol 1983 · PMID 6644819
Downloads1HHO.pdb / 1HHO.cif

pdb · 1HHO · retrieved 2026-09-18

01

Structures, callable

An agent pulls a real structure with method and resolution — no manual RCSB browsing.

02

Anchored to UniProt

A UniProt accession's cross-referenced PDB IDs resolve here directly — one protein, its structures.

03

Evidence travels with it

Every structure carries its resolution and primary citation, so trust is a fact, not a vibe.

PROTEINS & STRUCTURESAlphaFold DBPredicted structures for UniProt accessions, with confidence data.alphafold_structure
THE INTEGRATION

How Perslis integrates AlphaFold DB

AlphaFold DB (EMBL-EBI) hosts the DeepMind-predicted 3D structure for a UniProt protein, with per-residue confidence (pLDDT) and predicted aligned error (PAE) that quantify how reliable each region is.

On the console, alphafold_structure takes a UniProt accession and returns the model version, mean pLDDT, and download URLs for the CIF/PDB coordinates and the PAE.

Prediction is only useful with its confidence attached, so the instrument never returns a fold without its pLDDT. A high score (hemoglobin's is 98.06) says the model is reliable; a low one is a warning the console passes through unedited rather than hiding.

Because AlphaFold keys on the same UniProt accession as everything else on the protein bench, a predicted model and an experimental PDB structure line up for one protein automatically — the anchor is what lets the console show measured and predicted side by side.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · alphafold_structure("P69905")
ProteinHBA_HUMAN (Hemoglobin subunit alpha)
Model versionv6
Mean pLDDT98.06 (very high confidence)
CoordinatesAF-P69905-F1-model_v6.cif / .pdb
PAEpredicted_aligned_error_v6.json

alphafold · P69905 · retrieved 2026-09-18

01

Prediction with a trust score

Every model returns its pLDDT — an agent weighs the fold instead of assuming it.

02

Predicted beside measured

The shared UniProt accession lines an AlphaFold model up against its experimental PDB structure.

03

Low confidence is not hidden

A weak pLDDT is passed through verbatim — the console reports uncertainty, it does not smooth it away.

PROTEINS & STRUCTURESInterProProtein families, domains and sites — and an honest zero when nothing matches.interpro_domains · interpro_entry
THE INTEGRATION

How Perslis integrates InterPro

InterPro (EMBL-EBI) integrates a dozen protein-signature databases into one classification — families, domains, homologous superfamilies and sites — each with a stable InterPro accession.

On the console, interpro_domains returns the entries matching a UniProt protein, and interpro_entry resolves a single InterPro accession. A count of zero is a real result, reported as such.

The honest zero is the whole point. A model asked what domains a protein has will, unprompted, offer a confident-sounding list. InterPro replaces that guess with a lookup: hemoglobin returns the Globin domain and its superfamilies; an unclassified protein returns nothing, and the console reports nothing rather than filling the silence.

Keyed on the UniProt accession, InterPro completes the protein picture the bench builds — sequence and disease from UniProt, structure from PDB and AlphaFold, and composition from InterPro, all for one anchor.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · interpro_domains("P69905")
Returned6 entries (honest count)
DomainIPR000971 · Globin
FamilyIPR002338 · Hemoglobin, alpha-type
SuperfamilyIPR009050 · Globin-like superfamily
FamilyIPR050056 · Hemoglobin and related oxygen transporters
No matchwould return 0 — not a fabricated domain

interpro · IPR000971 · retrieved 2026-09-18

01

Composition, callable

An agent gets a protein's real domains and families, pinned to InterPro IDs — not a plausible guess.

02

Completes the protein picture

Composition joins sequence, structure and prediction on the shared UniProt anchor.

03

An honest zero

No match returns zero. The instrument is designed so the empty answer is a feature, not a hidden failure.

04
INSTRUMENT BENCH

Genomes & Variants

5 services
GENOMES & VARIANTSEnsemblGenes, coordinates, transcripts, FASTA sequence, cross-references.ensembl_gene · ensembl_sequence · ensembl_xrefs
THE INTEGRATION

How Perslis integrates Ensembl

Ensembl (EMBL-EBI) is the annotated reference genome for human and other species — genes, transcripts, coordinates and cross-references to protein and variant databases, all on a stated assembly like GRCh38.

On the console, ensembl_gene resolves a symbol to its record, ensembl_sequence returns FASTA, and ensembl_xrefs lists cross-references — the links that carry a gene to its protein and its variants.

Ensembl is the genome anchor, the way UniProt is the protein anchor. A gene resolved here — HBA1 to ENSG00000206172 on GRCh38 — carries the assembly with it, so a coordinate always means one place, and the cross-references hand off cleanly to VEP for variant effects, dbSNP for known variants, and NCBI for sequence records.

Every field is pinned: the gene ID, the region, the canonical transcript, the assembly. A genomic statement on the console names its coordinate system, because a position without an assembly is a position that could be wrong.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · ensembl_gene("HBA1")
Gene IDENSG00000206172
SymbolHBA1 (hemoglobin subunit alpha 1)
Biotypeprotein_coding
Region16:176,660–177,527 (+)
AssemblyGRCh38
Canonical transcriptENST00000320868.9

ensembl · ENSG00000206172 · retrieved 2026-09-18

01

The genome anchor

A gene ID keys straight to VEP, dbSNP and NCBI — one gene, its variants and sequence.

02

Coordinates, callable

An agent resolves a symbol to a position and sequence without opening the genome browser.

03

Always names its assembly

Every coordinate carries GRCh38 (or whichever build) — a position on the console is never ambiguous.

GENOMES & VARIANTSEnsembl VEPPredicted variant consequences from rsID or HGVS. Research use only.vep_consequences
THE INTEGRATION

How Perslis integrates Ensembl VEP

Ensembl VEP annotates a variant with its predicted molecular consequences across all overlapping transcripts, the most severe consequence, and pathogenicity predictions from SIFT and PolyPhen.

On the console, vep_consequences accepts an rsID or HGVS notation and returns per-transcript consequence terms, impact levels, and SIFT/PolyPhen scores — a prediction, clearly labelled as such.

VEP is prediction, and the console keeps that boundary sharp. It returns SIFT and PolyPhen scores as computed estimates, distinct from ClinVar's expert clinical classification — an agent sees the algorithmic call and the curated call as two different kinds of evidence, never merged into one confident verdict.

Keyed to the same coordinates as dbSNP and Ensembl, VEP completes the variant triangle: dbSNP says the variant exists, VEP predicts its effect, ClinVar reports what clinicians concluded. Every call is stamped research use only.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · vep_consequences("rs6025") — Factor V Leiden
Most severe consequencemissense_variant
Gene / transcriptF5 · ENST00000367796
ImpactMODERATE
SIFTdeleterious (0)
PolyPhenprobably_damaging (0.936)
AssemblyGRCh38 · 1:169,549,811

ensembl vep · rs6025 · retrieved 2026-09-18 · research use only

01

Consequence, callable

An agent predicts a variant's effect per transcript from an rsID — no VEP web form.

02

Prediction kept separate

SIFT/PolyPhen scores stay distinct from ClinVar's clinical classification — two kinds of evidence, not one.

03

Labelled research-use

Every result is stamped research use only, so a prediction is never mistaken for medical advice.

GENOMES & VARIANTSNCBINucleotide and protein records — FASTA and GenBank.ncbi_sequence_search · ncbi_fetch_sequence
THE INTEGRATION

How Perslis integrates NCBI

NCBI hosts the primary nucleotide (nuccore) and protein sequence databases — RefSeq curated records, GenBank submissions, RefSeqGene entries — the authoritative store for a sequence and its provenance.

On the console, ncbi_sequence_search queries either database and returns accessions with titles and lengths; ncbi_fetch_sequence retrieves the FASTA or GenBank record for one.

NCBI is where a raw sequence keeps its identity. A search for hemoglobin's gene returns NG_059186.1 — a RefSeqGene with an accession, a length, an organism — not an anonymous string of bases a model could have hallucinated. The accession is the receipt.

It sits beside Ensembl, which anchors the annotated genome, and UniProt, which anchors the protein: NCBI provides the underlying sequence records that back both, so a nucleotide or protein sequence on the console can always be traced to a submitted, accessioned entry.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · ncbi_sequence_search("HBA1 human mRNA RefSeq", db=nuccore)
Top accessionNG_059186.1
TitleHomo sapiens hemoglobin subunit alpha 1 (HBA1), RefSeqGene
LocusLRG_1225, chromosome 16
Length7,872 bp
Databasesnuccore (nucleotide) · protein
Fetchncbi_fetch_sequence → FASTA / GenBank

ncbi nuccore · NG_059186.1 · retrieved 2026-09-18

01

Sequence, callable

An agent searches and fetches nucleotide or protein records without EDirect scripting.

02

Backs the genome and protein

NCBI's accessioned records underpin the Ensembl and UniProt anchors on the same console.

03

A sequence keeps its receipt

Every FASTA carries an accession — a sequence on the console is never free-floating or invented.

GENOMES & VARIANTSClinVarVariant classifications and review status, reported verbatim — never re-graded.clinvar_search · clinvar_variant
THE INTEGRATION

How Perslis integrates ClinVar

ClinVar (NCBI) aggregates submitted interpretations of the clinical significance of genetic variants, with the review status that says how much scrutiny each classification received.

On the console, clinvar_search finds variants and clinvar_variant returns one — the classification, review status, associated condition and canonical SPDI — each field carried through untouched.

This is the sharpest edge of the whole runtime. A language model asked to judge a variant will happily produce a clinical opinion — which, for a real patient's variant, is dangerous. The console forbids it: ClinVar's classification passes through verbatim, review status attached, and the model is structurally barred from upgrading, downgrading or summarising it into a verdict.

ClinVar's clinical call sits beside VEP's algorithmic prediction and dbSNP's record so an agent sees them as separate evidence — the curated human classification never blurred with a computed guess. Every result is stamped research use only, not medical advice.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · clinvar_variant("17677") — the BRCA1 5382insC founder variant
VariantNM_007294.4(BRCA1):c.5266dup (p.Gln1756fs)
GeneBRCA1
Clinical significancePathogenic (reported verbatim)
Review statusreviewed by expert panel
ConditionBreast-ovarian cancer, familial, susceptibility to, 1
Variation IDVCV000017677.174

clinvar · 17677 · retrieved 2026-09-19 · research use only

01

Verbatim, by construction

Pathogenic stays Pathogenic. The model is structurally barred from re-grading a clinical classification.

02

Clinical kept separate

ClinVar's curated call never blurs with VEP's prediction — an agent sees two distinct kinds of evidence.

03

Review status attached

Every classification carries how much scrutiny it received — significance is never shown without its confidence.

GENOMES & VARIANTSdbSNPrsID records — location, genes, alleles.dbsnp_variant
THE INTEGRATION

How Perslis integrates dbSNP

dbSNP (NCBI) is the reference archive of short genetic variation — SNPs and small indels — each with a stable rsID, mapped location, alleles, functional class and any aggregated clinical significance.

On the console, dbsnp_variant takes an rsID (with or without the 'rs' prefix) and returns its position on a named assembly, its genes, alleles and the HGVS strings that other instruments consume.

dbSNP is the front door of the variant bench. An rsID resolved here — rs6025, Factor V Leiden, in the F5 gene at 1:169,549,811 — carries the assembly and the HGVS notations, which hand straight to VEP for predicted effect and to ClinVar for clinical classification. One identifier, three instruments.

Where dbSNP aggregates clinical significance, it is reported verbatim — the same never-re-graded rule ClinVar follows — and stamped research use only. A variant on the console always knows which genome build its coordinates belong to.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · dbsnp_variant("rs6025") — Factor V Leiden
Location1:169,549,811 (GRCh38)
GeneF5
AllelesC / A / G / T
Function classmissense_variant
HGVS (protein)NP_000121.2:p.Arg534Gln
Clinical significancepathogenic, risk-factor, … (verbatim)

dbsnp · rs6025 · retrieved 2026-09-18 · research use only

01

The variant front door

One rsID resolves to coordinates and HGVS that VEP and ClinVar consume directly.

02

Callable, assembly-stamped

An agent resolves a variant to a position that always names its genome build — no ambiguous coordinates.

03

Significance verbatim

Aggregated clinical significance is reported as-is, research use only — never re-graded by a model.

05
INSTRUMENT BENCH

Chemistry

2 services
CHEMISTRYPubChemCompounds — formula, weight, SMILES, InChIKey.pubchem_compound · pubchem_compound_by_cid
THE INTEGRATION

How Perslis integrates PubChem

PubChem (NCBI) is the largest open chemistry database — compounds with computed and curated properties: formula, molecular weight, SMILES, InChIKey, XLogP and more, each under a stable CID.

On the console, pubchem_compound resolves a name to its record and pubchem_compound_by_cid fetches one by CID — the canonical identity every downstream chemistry step keys on.

SMILES strings are exactly where a language model's confidence outruns its correctness — a plausible-looking structure can be subtly wrong. PubChem replaces generation with resolution: ask for aspirin and the console returns CID 2244 with the real InChIKey, not a string a model assembled. The identity is looked up, never invented.

That verified identity is the key for the chemistry bench: the same compound resolves into ChEMBL for measured bioactivity and into KEGG and Reactome for the pathways it acts on — one molecule, correctly identified, across four databases.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · pubchem_compound("aspirin")
CID2244
Molecular formulaC9H8O4
Molecular weight180.16
IUPAC name2-acetyloxybenzoic acid
Canonical SMILESCC(=O)OC1=CC=CC=C1C(=O)O
InChIKeyBSYNRYMUTXBXSQ-UHFFFAOYSA-N

pubchem · CID 2244 · retrieved 2026-09-18

01

Exact identity, callable

An agent gets a real CID, SMILES and InChIKey — no hand-written structure to be subtly wrong.

02

The key for chemistry

The same compound resolves into ChEMBL, KEGG and Reactome — one molecule across four databases.

03

Looked up, never invented

Structure is resolved from PubChem, so a SMILES on the console is the real one, pinned to CID.

CHEMISTRYChEMBLMeasured bioactivities — IC50, Ki — against biological targets.chembl_search · chembl_bioactivity
THE INTEGRATION

How Perslis integrates ChEMBL

ChEMBL (EMBL-EBI) is the manually-curated database of bioactive molecules — structures, calculated properties, clinical development phase, and measured activities (IC50, Ki, EC50) against biological targets, drawn from the literature.

On the console, chembl_search finds molecules and chembl_bioactivity returns the measured activities for one — each an experimental value with a source, not an estimate.

A potency number is the kind of specific-sounding fact a language model invents most convincingly. ChEMBL makes that unnecessary: aspirin resolves to CHEMBL25 with its real molecular weight and max clinical phase, and any activity value comes back as a measured datapoint tied to its assay — never a plausible figure with no experiment behind it.

ChEMBL joins the chemistry bench through the compound's identity: PubChem fixes what the molecule is, ChEMBL says what it does, and its target links reach back to UniProt proteins — the same anchor the structure bench is built on.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · chembl_search("aspirin")
ChEMBL IDCHEMBL25
Preferred nameASPIRIN
Max clinical phase4.0 (approved)
Molecule typeSmall molecule
Full MW180.16
AlogP1.31

chembl · CHEMBL25 · retrieved 2026-09-18

01

Measured, callable

An agent quotes a real IC50 or Ki tied to an assay — not a confident-sounding invented number.

02

Identity to activity

PubChem fixes the molecule; ChEMBL says what it does; targets link back to UniProt proteins.

03

Phase and provenance

Clinical phase and source travel with every molecule, so 'approved' is a fact you can check.

06
INSTRUMENT BENCH

Pathways

2 services
PATHWAYSKEGGPathways, compounds, enzymes, diseases, drugs.kegg_search · kegg_entry
THE INTEGRATION

How Perslis integrates KEGG

KEGG (Kyoto Encyclopedia of Genes and Genomes) links genomic and molecular information to higher-order functions — pathway maps and the enzymes, compounds, diseases and drugs that populate them, each with a stable KEGG identifier.

On the console, kegg_search finds entries and kegg_entry fetches the full flat-file record for one — a pathway with its gene list, compounds, modules and associated drugs.

KEGG turns an isolated molecule into a position in a process. A compound identified in PubChem or a target from ChEMBL can be placed on a KEGG pathway — glycolysis, hsa00010, with its enzyme genes, its compounds from glucose to pyruvate, and the drugs (mitapivat, etavopivat) that act on it — every element carrying a KEGG ID.

KEGG shares the pathways bench with Reactome; the two curate biological process differently, so the console offers both rather than collapsing them into one view. Every entry is pinned, so a pathway claim resolves to a KEGG record.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · kegg_entry("hsa00010")
PathwayGlycolysis / Gluconeogenesis (Homo sapiens)
ClassMetabolism; Carbohydrate metabolism
ModulesM00001 Glycolysis, M00002 core, M00003 gluconeogenesis
Example compoundsC00031 D-Glucose → C00022 Pyruvate
Associated drugsD11408 Mitapivat, D12362 Etavopivat

kegg · hsa00010 · retrieved 2026-09-18

01

Context, callable

An agent places a molecule in a pathway with its genes, compounds and drugs — pinned to KEGG IDs.

02

Chemistry meets biology

A PubChem compound or ChEMBL target lands on the pathway it acts in — one molecule, its machinery.

03

Two pathway views, not one

KEGG and Reactome both answer, so their different curation is a choice the console preserves.

PATHWAYSReactomeCurated pathways and reactions with stable identifiers.reactome_search
THE INTEGRATION

How Perslis integrates Reactome

Reactome is an open, manually-curated and peer-reviewed database of human pathways and reactions — molecular events organised into a navigable hierarchy, each with a stable identifier like R-HSA-70171.

On the console, reactome_search queries pathways, reactions and proteins by name or species and returns entries with their stable IDs — the anchor a pathway claim resolves against.

Reactome and KEGG both map biological process, but they curate it differently — reaction-level detail versus pathway maps — so the console keeps both rather than picking a winner. A question about glycolysis returns Reactome's R-HSA-70171 and KEGG's hsa00010, and an agent sees two independent curations of the same biology.

Because Reactome's proteins carry UniProt accessions, a pathway found here reaches back to the protein bench — the same anchor that ties structure, prediction and domains together, now placing a protein inside the reactions it participates in.

Open the source ↗
REAL, SOURCE-PINNED RESULTLive pull · reactome_search("glycolysis")
Top pathwayGlycolysis
Stable IDR-HSA-70171
TypePathway
SpeciesHomo sapiens
RelatedR-HSA-70326 Glucose metabolism

reactome · R-HSA-70171 · retrieved 2026-09-18

01

Curated pathways, callable

An agent reaches peer-reviewed reactions pinned to stable R-HSA identifiers — no manual browsing.

02

A second curated view

Reactome answers alongside KEGG, so two independent curations of a pathway are both on the table.

03

Reaches back to proteins

Reactome proteins carry UniProt accessions, tying a pathway to the structure bench's anchor.

07
NATIVE BENCH

Native & WetHands

11 instruments
COMPUTECompute benchPython beside the data lanes for sequence work and cheminformatics.run_python
THE NATIVE INSTRUMENT

Compute on what the instruments retrieved.

Biopython, pandas, SciPy, NumPy and RDKit sit beside the source lanes, so sequence and molecule work can run without copying evidence into another application. It is labeled honestly: process isolation, not a security sandbox.

TOOL CONTRACTrun_python
InputSource-pinned records and code
OutputCaptured result, parameters and job state
BoundaryProcess isolation; not a security sandbox
01

Evidence stays attached

Compute beside the retrieved record instead of starting a disconnected notebook.

02

Parameters remain visible

Inputs, outputs and settings stay with the research state.

03

Unknown is valid

A failed computation cannot be promoted into a scientific fact.

ACTUATIONWetHandsTen instruments for durable jobs and workspace-contained files.wetware_job_start · wetware_job_status · wetware_job_output · wetware_job_cancel · wetware_jobs · wetware_write_file · wetware_read_file · wetware_move · wetware_remove · list_workspace
THE WET LANE

Experiments that outlive the session.

WetHands runs managed background work for up to 24 hours with saved progress, monitoring and cancellation. Its file operations stay inside the assigned workspace. The same job contract is the integration surface for LIMS, electronic lab notebooks and instrument automation APIs; no autonomous laboratory is claimed.

TEN-INSTRUMENT LANEJobs + outputs + workspace files
Lifecyclestart · status · output · cancel · list
Fileswrite · read · move · remove · list
RetentionSaved progress survives restarts
01

Durable work

Longer jobs continue after the conversation closes.

02

Controlled files

Every file operation remains inside the assigned workspace.

03

Review before reliance

Outputs return to the same evidence gate before entering the record.

Services are queried through their public interfaces under the source contract. Names and marks belong to their owners; no partnership or endorsement is implied. Variant and clinical instruments are for research use only.

EXPLORE THE DETAILS

Every part of the work, connected.

01

Scientific databases

Search the instrument wall: PubMed, PMC, arXiv, UniProt, RCSB PDB, AlphaFold, Ensembl, ClinVar, PubChem, ChEMBL, KEGG and more. Each service keeps its own identity and source record.

Explore Scientific databases
02

APIs

Connect supported scientific services to your workflow and review results alongside their sources. Confirm availability and access with the team before relying on an integration.

Explore APIs
03

Instruments

Connect computational tools and explore WetHands for lab workflows. Device access and physical execution require a configured integration and an agreed scope.

Explore Instruments
04

MCP

The research server exposes scientific tools through Model Context Protocol. Compatible assistants can query the same source-pinned instrument surface.

Explore MCP
05

Files

Keep exported topic packs, records and research notes with your project. Local evidence remains inspectable when a live source is unavailable; fresh retrieval still needs its service.

Explore Files
06

Legacy systems

Extend a research workflow to existing software through the Perslis legacy-systems work. Start with a bounded task and validate the connection before relying on it.

Explore Legacy systems
SEE IT IN ACTION

See the instruments work together.

A recorded research pass across connected scientific sources.

Recorded on a research prototype. The recording shows the scope demonstrated at that time.

Explore the research floor yourself →
PERSLIS / SCIENCERECORDED DEMO
A recorded research pass across connected scientific sources.
ONE CONTINUOUS THREAD

From a question to a record you can inspect.

01Ask
02Retrieve
03Pin
04Verify
05Retain
See how it works →
BRING LOIS A REAL QUESTION

Let your next discovery start here.

Start with a target, a dataset or a research question. Inspect what the evidence supports, what it rejects and what remains unknown.

Start a research pilot →

Everything Perslis