Ontech ICTM platform

IT infrastructure management software for servers, hardware and services

Watch server health, hardware alarms and service uptime without installing agents — and keep maintenance planned rather than reactive.

What is IT infrastructure management?

IT infrastructure management is the ongoing work of keeping servers, hardware, platforms and the services running on them available, healthy and maintained. Ontech ICTM supports it with agentless server monitoring over SSH, Prometheus metrics, hardware alarms read from server management controllers through Redfish or IPMI, HTTP uptime checks, container discovery and maintenance scheduling.

Infrastructure problems rarely start as outages. A disk fills, a power supply fails in a redundant pair, a service stops restarting, a certificate becomes invalid. Infrastructure management is about seeing those signals early, acting on them in a controlled way and keeping a record of what was done.

ICTM gathers the signals through methods most environments already support. It connects to servers over SSH to collect operating-system metrics and service status, reads history from Prometheus node_exporter where you run it, and polls the baseboard management controllers built into server hardware — Dell iDRAC, HPE iLO, Lenovo XCC, Cisco CIMC and Supermicro — through Redfish or IPMI.

Alarms are evaluated against rules you control and can be acknowledged, suppressed, escalated or resolved, while maintenance schedules and logs sit alongside the monitoring so planned work and unplanned incidents are visible together.

What ICTM does for infrastructure management

Agentless server monitoring

Collect CPU, memory, disk and network-interface figures over SSH, along with running and failed services, top processes, failed logins and sudo events, open ports, Docker containers and web-server logs — no agent to install.

System metrics from Prometheus

Pull CPU, memory, disk and network history from Prometheus node_exporter so trends are visible alongside the rest of your infrastructure records.

Hardware alarms from server controllers

Poll Dell iDRAC, HPE iLO, Lenovo XCC, Cisco CIMC and Supermicro controllers through Redfish or IPMI on a schedule (every five minutes by default), apply threshold rules, and acknowledge, suppress, escalate or auto-resolve alarms.

HTTP uptime and application checks

Schedule HTTP checks with an expected status code and SSL verification, and combine them with SSH health checks and metric collection for each application you depend on.

Grafana integration

Connect an existing Grafana to see its health, data sources and dashboards, and embed Grafana panels inside ICTM.

Container discovery

Find the Docker containers running on a server over SSH, import them as records and read their logs, so containerised services appear in the same inventory as everything else.

Browser-based remote terminal

Run SSH commands against a managed server from the browser, keeping routine checks inside the platform.

Maintenance schedules and calendar

Plan maintenance windows, log completed work and see upcoming activity on a calendar next to the monitoring data.

Rule-based risk scoring

Predictive analytics scores systems against CPU, memory and disk thresholds and equipment age using transparent rules, highlighting infrastructure that deserves attention before it causes an incident.

Cloud resource and cost records

Record cloud providers and resources and roll up the costs you record against them, giving one register for on-premises and cloud infrastructure.

ICTM modules: Infrastructure · Systems · Systems Monitoring · Server Monitoring · System Metrics · Hardware Alarms · Alarm Rules · HTTP Monitors · Grafana Dashboards · Remote Terminal · Containers · Cloud · Maintenance Calendar · Predictive Analytics

Use cases

Data-centre hardware health

Catch temperature, fan and power-supply faults on rack servers through their management controllers before redundancy runs out.

Customer-facing service uptime

Check that web applications and portals respond with the expected status and a valid certificate, on a schedule.

Branch and regional servers

Monitor servers in other offices over SSH without installing and maintaining monitoring agents on each one.

Planned maintenance

Schedule patching and hardware work, record the outcome and keep the history for audits and post-incident reviews.

Consolidating monitoring views

Bring Prometheus metrics and Grafana dashboards you already run into the same place as assets, alarms and incidents.

Benefits

Earlier warning

Hardware alarms and threshold rules surface problems while redundancy or headroom still protects users.

Low deployment effort

SSH, Redfish, IPMI and Prometheus are standard, so there is no agent estate to install, update and secure.

Context for every alert

Alarms relate to asset records, owners and dependencies, so the right people know what is affected.

A record of what was done

Acknowledgements, escalations and maintenance logs create an operational history that supports audits and reviews.

Frequently asked questions

Does ICTM need agents installed on servers?

No. ICTM collects server metrics and status over SSH, reads hardware health from the server's management controller through Redfish or IPMI, and can pull metrics from Prometheus node_exporter if you already run it. You provide credentials for the systems you want monitored; nothing needs to be installed on them.

Which server hardware can ICTM read alarms from?

ICTM polls baseboard management controllers that support Redfish or IPMI, including Dell iDRAC, HPE iLO, Lenovo XCC, Cisco CIMC and Supermicro. Polling runs every five minutes by default, and alarm rules decide which readings raise an alarm and at what severity.

Can ICTM monitor websites and web applications?

Yes. HTTP monitors check a URL on a schedule, compare the response with the status code you expect and verify the SSL certificate. For applications on servers you manage, ICTM can add SSH health checks and metric collection so you see both the service and the host behind it.

How does ICTM flag infrastructure at risk?

Predictive analytics in ICTM is rule-based: it scores systems against CPU, memory and disk thresholds and equipment age, together with factors from other modules. Because the rules are transparent, your team can see exactly why a system was flagged and adjust maintenance plans accordingly.

Does ICTM work with Prometheus and Grafana?

Yes. ICTM reads CPU, memory, disk and network history from Prometheus node_exporter, and connects to an existing Grafana to show its health, data sources and dashboards and to embed panels. You keep the tools you have and view them alongside ICTM's own asset and alarm data.

Can ICTM manage cloud and container infrastructure?

ICTM finds Docker containers on a server over SSH, imports them and shows their logs, and it keeps records of cloud providers, resources and the costs you record. It is an inventory and monitoring layer: it does not provision cloud resources or control container orchestration.

See infrastructure management in Ontech ICTM

Book a walkthrough with the Ontech team, or start a free trial and explore the platform yourself.