Showing posts with label Popular. Show all posts
Showing posts with label Popular. Show all posts

Saturday, March 23, 2013

A day in the life of: The Systems Administrator


Systems Administrator:


A Systems admin can go by many names: Server Admin, NT Administrator, Systems Engineer, Unix/Linux admin, Network Administrator, etc. Refer to job descriptions to more clearly understand the company’s definition for their environment and the name they assign to this role. In the end, an IT admin supports infrastructure systems and devices.

In this “day in the life”, I'll focus on a Windows server administrator.

I would also like to mention that in my career, I have specialized in helping companies where the executive management team has reached their boiling point with their IT Infra teams. They reached out to me for help come in and assess the environment, develop a design and action plan then put things into place. The approach I have fine-tuned over the years may be the subject of a series of other posts if there is interest.

Windows is by far the most common server operating system you will see in the enterprise. The Linux/Unix admins have similar workdays however there are some differences that I will cover in another post. Network Admin (for switches and routers) is yet another posting.

Group Role:


The Windows server support group, along with the network group, is the foundation of the entire business. The physical and virtual server that the team installs, configures, secures and essentially keeps ‘alive’ and running at peak performance hosts the critical application which allows the company to do business.

In addition to the keeping the servers running and secured, the team typically supports applications such as: Email, Active Directory, ADFS, File & Print, Group Policy (automation and policy enforcement), Systems Management tools (IE SCCM, MDM, SIEM, IDMS), SharePoint, Corporate Antivirus management and  Helpdesk ticket systems.

Start of day: 


1.      Wake up, rub your eyes and grab your smartphone to check your work email if any alerts or critical requests have been sent to you.
·         No I’m not kidding… server admins are a special breed who really cares about their systems and are always keeping an eye on things 

2.      If any alerts or emails need immediate attention (IE an alert a server’s CPU has been running at 99% or a drive is near full capacity)
·         jump on your laptop and connect to the VPN
·         Remotely connect to the server in question and remedy the problem
·         If a ticket wasn’t created for the alert, create one. Update accordingly and close the ticket
·         Log off 

3.      Go through your morning routine then head to work.
·         Note: It’s always a good idea to arrive early to get a jump on any issues before users are calling about some server issue preventing them from doing their work. 

4.      Arrive at work and check the ‘state of the union’ by checking in with the NOC (if you have one), check the helpdesk tickets assigned to you and your group & the monitoring dashboard(s).
·         If any critical issues are open, crack open the laptop and remediate the problem.
·         Update and close the ticket related to the the issue you resolved. If a ticket doesn’t exist, create one. 

5.      If all is good in the world of servers, head for your morning coffee.
·         Coffee in the morning is a perfect time to interface with the colleagues within your IT group as well as other groups.
o Collaboration strengthens IT as a whole and gives you some insight if there are some other issues going where you may be able to help.
·         Coffee time is also a great time to talk to end users (non-IT) and find out how things are going
o Listen to what users say closely and read between the lines. If you pick up on any pain points like an app is slow today or challenges that some new software could increase their productivity or quality of life…
o Make a mental note and look into it. If its new software solutions, discuss it with your manager. If you fix a problem for the user, follow it up in an email to let them know otherwise they think the problem just went away.
·         Pet peeve note: Careful with extended coffee runs.
o Many managers interpret long coffee breaks as a negative unless they are aware of your 'recon' motive
o Users who get the impression IT is always on break or arent busy is bad for IT PR.
o Find the right balance of showing time to listen to people and lingering too long

6.      Return to your desk and again, check on the ‘state of the union’ (Dashboard and tickets) 

7.      Begin your normal work assignments
·         This may vary depending on how your manager assigns duties and priorities. What I found works well is in this following order:
1) Resolve any critical incident tickets  (something broken)
2) Resolve any critical requests (something needed)
3) Proactive health checks (20% of your day)
·        Each engineer should have a list of systems and app to review on a daily, weekly and monthly basis to resolve problems before they become a risk to an outage.
4) Resume resolving tickets (60% of your day)
5) Project work (20% of your day)
·        Some projects you may be dedicated 100% of your day for the duration of the project. This will be determined by your manager based on criticality and the project completion date 

8.      Throughout the day, expect interruptions. These will cause you to shift from one assignment to another.
·         System engineers juggle a lot of tasks concurrently during the day as all other IT groups rely on the infrastructure teams. Their critical events translate to a critical response by Infra groups in order for IT to respond as a whole to the ever changing requirements and needs of the business. 

9.      End of shift
·         Don’t expect to leave exactly on time. What we do as system admins doesn’t exactly finish on our schedule. It is part of the life we chose in supporting servers that don’t turn off. It is our responsibility to keep them and the business going.
·         Bring your laptop home with you every day. If you are not on call, one of your team members is and may need help. Great teams always back each other up.
·         Check the monitoring dashboards one last time before leaving 

10.  End of day

·         The end of the day doesn’t really end when your shift does as on-call rotations are common. Avoid becoming burnt out by keeping up with other hobbies outside of tech but definitely do not disconnect completely when going home for the day.


Server admins are in their position because they are never really disconnected (on-call or not). Part of this is because they become personally connected to their systems and the company. We take great pride in what systems we have built.  

Common tickets for System Administrators:

·         Email [spam, delivery problem, mailbox needed]
·         Need a new server for “xyz” project
·         Patch or software deployment needed to [servers, pc/laptop, users with “X” installed]
·         Mailbox or distribution list needed
·         Alert received for server “X”  [connectivity lost, cpu pinned @99%, free space is low, unauthorized access attempt, URL check failed, etc]
·         Data management [file restore, access request/removal, new share]
·         New solution or service project [Proof of concept, test, pilot, deploy]
·         Server or application performance
·         “How do I…” or how to questions 

Supplemental


Becoming a Systems Administrator

Server admins typically have their careers start within the helpdesk or NOC where a great amount of experience is gained, more so from the helpdesk. Many companies promote from within when a helpdesk technician shows the traits and technical initiative needed for server side support.

Skills and traits of a good admin

·         Integrity – First and foremost this trait is absolutely required.

·         Commitment – Servers can crash at any time and the business depends on IT to do whatever it takes to get services restored.
o   Taking part of an on-call rotation is very common
o   Helping the on-call person respond to an alert builds team unity
o   Seasoned engineers can all share a story of at least one time they had to work 24+ hours straight on restoring a server or recovering all the systems in the DR site (disaster recovery site). This should be a rarity but it will definitely happen in your career. This level of commitment is rewarded greatly, not just monetarily but also in the business relationships you develop with your executive management team.

·         Stay ahead of the curve
o   At this level of IT, you are the one who should have a solution to every challenge of at least know how to find it quickly. Stay ahead on your certifications and tech news

·         Vision
o   See the forest, not just the trees.
o   For example:
§  If every day you need to reboot a server to correct an issue; that is not a fix but a Band-Aid. Find the real problem so you no longer have to do this daily manual task.
§  The helpdesk keeps getting calls to set a new users homepage, desktop shortcuts and mapped network drives to match what the rest of their group has. Create a GPO (group policy object) to automatically create these settings and save the desktop group hundreds of hours over the next couple years.
§  An application owner or even a Microsoft support rep requests you to make a change the server that is having a problem. Will this cause a problem now or in the future? Will it violate any corporate standard or policy? Knowing the answers to these questions is critical.

·         Diagnostic skills
o   A professor I once had the pleasure of having once described diagnostic or troubleshooting skills as a talent that “you either got it or you don’t”. I agree with him however I do believe these skills need to be developed through use to become stronger.
§  Pay particular attention to how components or pieces come together. What dependencies are there between components and where are the failure points.
§  Identify what symptoms relate to different failure points. For example, if Internet Explorer displays a message “page cannot be displayed’; what’s wrong (if you can name only 1 or 2 things, it’s time for you to do some studying)

·         Ability to work with little supervision

o   Depends on your manager’s style however most good managers do not attempt to micromanage the systems engineers who are often relied upon to be subject matter experts (SMEs).

o   If you have built the confidence your manager has in you and judgment, they will normally loosen the reigns. If you find you are being managed more closely than your peers, evaluate how you can improve by using your peers as examples as well as speaking to your manager. In order to advance, your manager needs to really trust in you. 

Sorry this was such a long post folks but this job role manages a considerable amount of the overall IT presence within a company. The success of the server team has a direct impact upon the productivity of the business and should be a leader of innovation and security initiatives. Carreers as System Administrators can be very rewarding with talented engineers always in high demand.

 

Sunday, March 3, 2013

A day in the life of: Helpdesk Support

The Helpdesk


Group Role:
The helpdesk is THE face of IT. Users look to the helpdesk for answers to their problems as well as to help improve their productivity by installing new software or getting a new computer/laptop. As a helpdesk technician, you need to appear as always having the right answer and be completely trustworthy. After all, a user who hands over their laptop with all of their critical and confidential data needs to be confident you wont lose their files or look at the contents.
  • Hardware support:
    • Desktops and laptops from multiple vendors such as HP, Dell, IBM, and Apple. Replace failed components and issue new systems. 
    • Printers, scanners and sometimes basic mobile device support depending on the environment. Typically hardware support is minimal triage before calling a 3rd party for more advanced support (IE setup a new printer on the network, change toner and rollers but not take the thing apart)
  • Software support:
    • Microsoft Windows or Mac OS: Install, troubleshoot, patch, upgrade
    • Microsoft Office, Adobe, print devices: install, troubleshoot, patch, upgrade
    • Install and Other 3rd party software the user may require
    • Troubleshoot errors the user may see
      • Internet Explorer alerts or pages not loading
      • No network available (wired/wireless)
      • Unexpected crashes of an application or OS
  • Account support
    • Account creation and removal
      • May be the responsibility of the helpdesk or as organizations get larger, there is more segregation of duties. The task may be assigned to the security team or even be delegated to HR and created when someone is on boarded or off boarded.
    • Password resets, Unlock accounts
    • Add/Remove groups to Active Directory or update memberships to a group
      • Sometimes this is delegated to the Helpdesk or managed by the Wintel (server) team.
    • Computer accounts in Active Directory
      • New PCs are joined to the domain when issued to a user
      • Old PCs are removed from the domain when decommissioned.


Start of Day:
  1. Arrive early.
    • Coming in every day 15 minutes early shows your manager you are dependable and committed. It also helps you get the jump on any critical ticket that came in. Most executives come in early and leave late. Being the first one in translates into your becoming the go-to person for executives. This goes a long way.
  2. Check the helpdesk ticket system for any open tickets that are critical.
    • Pay close attention to not only the ticket request/problem but also to WHO put in the request. A medium request by your CFO warrants more than just a medium response.
    • Many managers allow their support teams to self assign tickets. If so, don't cherry pick the easy tickets. Be sure to take on the more difficult tickets if you are capable and become more valuable to the team.
    • If there are a number of tickets with the same issue, talk to the associated engineering team or application team if this is a broader issue occurring. [IE 20 people not able to login to Sharepoint]
  3. Check with the NOC and other teams on any open issues such as outages or planned maintenance.
  4. Phone support / Deskside Support
    • Some helpdesk groups assign specific technicians for phone support and others to deskside support. Often its a combination of both. Depending on how your department is setup, at this point of the day begin manning the phones and responding to ticket requests.
  5. Tickets
    • Most small to medium groups utilize some helpdesk system while large enterprises its a must for managers and support teams.
    • Follow your departments procedures on creating, assigning and updating tickets.
    • Always enter detailed information when updating a ticket. When your manager reviews your work, he/she can fully understand how well you are doing and what is going on.
End of day
    • IT positions never end on time so don't count on running out at exactly 5PM. Very often you may be in the middle of installing some software for a user or transferring their data to a new PC.
    • Not watching the clock and staying late are good ways to get noticed by your manager and builds their confidence in you.  (Note: CIOs and business leaders users also notice)
Common tickets to expect to be assigned:
  • Account is locked out
  • Issue (load) a computer or laptop for a new hire
  • Need software installed
  • Computer wont boot or computer crashed 
  • Cant access/send an email (Outlook)
  • Not connected to the wireless network
  • Office relocation - setup the user's PC in their new office or cube


Supplemental:
Deskside positions can lead to other opportunities such as Helpdesk manager, Server engineer, Network Engineer, Systems administrator, Active Directory Admin, and others.

Certifications often held by a support technician:
 Qualities as a hiring manager that I look for:
  1. Integrity
    • I must be able to trust the candidate to a certain ethical and professional level. IT positions come with a lot of responsibility and access to sensitive information. This trust must be earned not only from the manager, but the end users that IT supports.
  2. Diagnostics
    • No matter how much experience someone has, if they cant instinctively troubleshoot than fixing an executives laptop while he/she is pressing you for answers and immediate resolution.. lets just say it just wont end well.
    • Be aware of how to research issues and weed out the false or damaging resolution suggestions.
  3. Commitment
    • As an IT professional, problems will come up that you dont have the answer to...yet. There are those types of folks who hit the first obstacle and look to someone else to give them the answer. Others folks dont give up so easily and are committed to finding the solution. Note that there is a point where you should ask for help.Never asking for help is almost as bad as always asking for it.
    • Commitment also includes staying ahead of the curve on technology. Businesses look to IT for answers. If the business users know more about a new gadget or IT solution coming out than you do, it is a good indication your skills and knowledge are behind the curve.
  4. Interpersonal skills
    • Members of the department or team must get along. Personality conflicts causes friction which leads to bigger problems.
    • Having a group of people who always has each others back and are very collaborative leads to success for IT as a whole.
    • A person who can effectively communicate via email, phone or in-person with IT and non-IT personell is a plus. For the helpdesk technicians, this is critical.

A day in the life of: NOC agent

NOC: Network Operations Center

Group Role:
  • Monitor system health
  • Monitor network connections
  • Monitor batch job completion, job scheduling (rerun, pause, execute)
  • Tier 1 responses often documented in a 'run book'
  • Escalation of priority alerts to the appropriate support team
Note: A good place to get started in IT can be as a NOC agent

Start of day:
  1. Get a lot of coffee.
    • The NOC runs 24x7 and staring at a screen waiting for an alert can be a challenge.
    • Some NOCs run 12 hour shifts with agents working 3-4 days per week.
  2. Status transition
    • As each shift changes, status of open items need to be communicated to the agents coming in, such as:
      • Current outages and actions taken so far in the run book
      • Systems under maintenance and when to resume monitoring
      • Open requests (ie  start job #2456 @ 2:15AM)
      • Changes to on-call coverage
    • Confirm the information being handed off is correct (check the status board for open alarms)
  3. Monitor
    • Most NOCs have several dashboards that show the current state of the environment.
    • Alarms are usually very easy to recognize with big red indicators, warnings are yellow
    • A mailbox is usually used as well for alert messages, batch job failure or success notices and requests for batch jobs to be run/held.
  4. Alert!
    • An alert comes up. Follow the response checklist which may look like this:
      1. Confirm if the alert is a false positive
      2. Enter a ticket to record the event and assign to the NOC
      3. Perform any remediation steps documented for such an alert (if present)
      4. If the alert doesn't clear, follow the escalation procedures
        • P1 or Critical priority response is typically a voice hand-off to the on-call engineer who handles the system or application the alarm is for
        • Medium to low priority alarms, such as a QA sytem being offline are typically an email notice to the support team and ticket being assigned to them
      5. Update the ticket and reassign it to the appropriate support team/person.
  5. Batch Request
    • Users will request a batch job to run or be held very often. If one fails, re-running the job is the normal action taken.
    • Multiple failures: submit a ticket to the support team for remediation.
  6. Notices
    • Often engineers will notify the NOC to ignore alarms for a system as maintenance will be performed. Depending on the monitoring software, typically a NOC agent will put the system into 'maintenance mode' which suppresses alerts during a specified time window.
    • On-call coverage changes occur frequently. These changes to the default rotation should be noted and also communicated to the next shift.
  7. End of shift
    • Transition status of systems, maintenance windows and other details to the next shift.

End of Day.


Supplemental:
There is a lot of quiet time. NOC agents can best make use of this time towards training, independent study and certifications. A natural advancement is into one of the Infrastructure support teams.



Tuesday, February 26, 2013

Staying ahead of the curve

If you're not ahead of the curve then you are behind the curve

How to keep up with training with little or no cost?

Technology moves faster every day and many businesses have reduced or eliminated their training budgets yet as IT professionals, if we dont stay ahead we begin to loose our value.

Another way to look at it:
     If you stay ahead, you increase your value.

It's difficult reading technical books even for the most geeknerds (yes, I am one) but there are other options out there even if your employer wont flip the bill to send you for training. The effort you invest in your skills and knowledge will continue to pay dividends. You're value is in what you know and in how much the business can rely upon you.

So what can you with little or no cost?

  • (Free) Attend User group meetings in your area.
    • User groups are great ways to learn from peers in trends, what are some things that have worked out well or not so well.
    • Meetings usually have sponsors that give away some pretty cool stuff as door prizes.

  • (low cost) Purchase a Microsoft Technet subscription ($149 for standard)
    • Eval and beta versions of MS software for 12 months (use, play, learn)
    • Around 20 hours of e-learning
    • 24/7 online chat assistance
    • Refer to the full details on Microsoft's site

  • (low cost) Industry publications
    • Magazines (online or print) can help keep you in the loop of upcoming trends or solutions that can help the business units you support.
    • Example: I love WindowIT Pro Magazine each month I find at least something useful

  • (Free) Quick video training
    •  Streaming videos help engineers get a quick start on new topics
    • Some examples are included below
Microsoft System Center 2012

 
Installing VMware vCenter Server
 
 
Upgrading to Active Directory 2012
 

  • Finally, sorry to say it but we wont get away from reading. The last item is purchasing a detailed book. Pick these when it becomes apparent that you will need to become the resident expert on a software package or technology. Install the software in a home lab using your technet subscription.

Once again, please never stop learning. At times it's a struggle but it is an investment into your own career. If you fall behind don't be surprised when you are left behind.


Good luck and God bless